GLM 5.3 Flash aka Ox Alpha is now live in Command Code. 4x usage in GOAT plan. $40 credits. ~24K reqs. ~1.2B tokens. Available across all plans and API 🐐
GLM-5.3 Flash at a glance
Fast, affordable GLM coding with 1M context.
| Spec | GLM-5.3 Flash |
|---|---|
| Model ID | z-ai/glm-5.3-flash |
| Context window | 1.05M tokens |
| Released | August 26, 2026 |
| Intelligence Index | 41.8 |
| Coding Index | 71.5 |
| Input | $0.15 per 1M tokens |
| Output | $0.50 per 1M tokens |
| Cache read | $0.03 per 1M tokens |
| Effective input in an agent loop | $0.07 per 1M tokens |
| Plans | Go and above |
Specs, pricing and plans are from the GLM-5.3 Flash model page as of September 28, 2026, including any deal running that day, so they can differ from the launch pricing above. Intelligence and Coding Index scores are from Artificial Analysis.
Use GLM-5.3 Flash
1cmd --model z-ai/glm-5.3-flashOr type /model in a session and pick it. You can switch mid-session without losing context.
What real work costs on GLM-5.3 Flash
| Task | Tokens (in · out) | Cost |
|---|---|---|
| Quick lookup / one-liner | 8K · 1K | $0.0008 |
| Review a 500-line PR | 60K · 4K | $0.0060 |
| Fix a bug (agent loop) | 180K · 12K | $0.02 |
| Refactor a module | 320K · 20K | $0.03 |
| Full-repo agent run | 900K · 45K | $0.07 |
Try Command Code
The best coding agent for open models, in your terminal or on your desktop.
1npm i -g command-codeDesktop App · Docs · X · Discord
