We're increasing GLM 5.3 Flash usage to $60 of credits in Command Code, which means 6x usage on the GOAT plan. 🐐 GLM 5.3 Flash has a 1M token context window, handles text, images and video, and is built for fast coding and long-horizon agent tasks.
GLM-5.3 Flash at a glance
Fast, affordable GLM coding with 1M context.
| Spec | GLM-5.3 Flash |
|---|---|
| Model ID | z-ai/glm-5.3-flash |
| Context window | 1.05M tokens |
| Released | August 26, 2026 |
| Intelligence Index | 41.8 |
| Input | $0.15 per 1M tokens |
| Output | $0.50 per 1M tokens |
| Cache read | $0.03 per 1M tokens |
| Effective input in an agent loop | $0.07 per 1M tokens |
| Plans | Go and above |
Specs, pricing and plans are from the GLM-5.3 Flash model page as of October 6, 2026, including any deal running that day, so they can differ from the launch pricing above. Intelligence and Coding Index scores are from Artificial Analysis.
Use GLM-5.3 Flash
1cmd --model z-ai/glm-5.3-flashOr type /model in a session and pick it. You can switch mid-session without losing context.
What real work costs on GLM-5.3 Flash
| Task | Tokens (in · out) | Cost |
|---|---|---|
| Quick lookup / one-liner | 8K · 1K | $0.0008 |
| Review a 500-line PR | 60K · 4K | $0.0060 |
| Fix a bug (agent loop) | 180K · 12K | $0.02 |
| Refactor a module | 320K · 20K | $0.03 |
| Full-repo agent run | 900K · 45K | $0.07 |
Try Command Code
The best coding agent for open models, in your terminal or on your desktop.
1npm i -g command-codeDesktop App · Docs · X · Discord
