GLM-5.3 Flash
z-ai/glm-5.3-flash
fast, affordable GLM coding with 1M context.
cmd --model z-ai/glm-5.3-flash
Intelligence index
not yet scored
Output speed
not yet scored
Input
$0.15 /M
Output
$0.50 /M
Cache read
$0.03 /M
Agent-loop cost
$0.07 /M in
Context window
1.05M tokens
Released
—
Modalities
→
vs. the lineup
GLM-5.3 Flash beside its stablemates and nearest rivals. The ◆ marks the best value in each column across every row shown.
pin a rival:
| Model | Intelligence | Coding | Speed | Input $/M | Output $/M | Blended $/M | Context |
|---|---|---|---|---|---|---|---|
| GLM-5.3 | 59.5◆ | 74.8◆ | 79.6◆ | $1.40◆ | $4.40◆ | $2.15◆ | 1M◆ |
| GLM-5.2 | 52.6◆ | 68.8◆ | 69◆ | $1.40◆ | $4.40◆ | $2.15◆ | 1M◆ |
| GLM-5.1 | 41◆ | 55.8◆ | 28.2◆ | $1.40◆ | $4.40◆ | $2.15◆ | 200K◆ |
| GLM-5 | 40.6◆ | —◆ | —◆ | $1◆ | $3.20◆ | $1.55◆ | 200K◆ |
| GLM-5.3 Flash ◆ | —◆ | —◆ | —◆ | $0.15◆ | $0.50◆ | $0.24◆ | 1.05M◆ |
| GLM-5.2 Fast | —◆ | —◆ | —◆ | $3◆ | $10.25◆ | $4.81◆ | 1M◆ |
| DeepSeek V4 Pro (latest)pinned | 53.2◆ | 68.8◆ | 75.3◆ | $0.66◆ | $1.98◆ | $0.99◆ | 1M◆ |
| DeepSeek V4 Flash (latest)pinned | 52◆ | 69.1◆ | 114.6◆ | $0.22◆ | $0.66◆ | $0.33◆ | 1M◆ |
usage calculator
How far a month of credits goes on GLM-5.3 Flash.
Input tokensfresh prompt
800Output tokensmodel reply
180Cache read tokensre-read context
50Kcost / request $0.0017 · in $0.15 · out $0.50 · cache $0.03 per M
fresh input 7%output 5%cache reads 88%
Requests / 30 days
5.8K
$10 credits ÷ $0.0017 per request
~1.2K quick fixes~234 bug fixes~39 feature PRs
what real work costs
Real coding tasks priced end to end on GLM-5.3 Flash, from a quick lookup to a full-repo agent run.
One agent task
$0.02
180K in at 75% cache hit, 12K out
What you pay
$0.09 /M
all-in across every token that task touched
Sticker input
$0.15 /M
cache reads bill at $0.03 /M instead
| Task | Tokens in · out | GLM-5.3 Flash | Muse Spark 1.2 Contributor | Claude Haiku 4.5 |
|---|---|---|---|---|
| Quick lookup / one-liner | 8K · 1K | $0.0008 | $0.0003 | $0.0065 |
| Review a 500-line PR | 60K · 4K | $0.0060 | $0.0027 | $0.04 |
| Fix a bug (agent loop) | 180K · 12K | $0.02 | $0.0072 | $0.12 |
| Refactor a module | 320K · 20K | $0.03 | $0.01 | $0.21 |
| Full-repo agent run | 900K · 45K | $0.07 | $0.03 | $0.49 |
frequently asked
How much does GLM-5.3 Flash cost?
$0.15/M input and $0.50/M output, cache reads $0.03/M. In an agent loop most input is cache-read, so the effective input rate is about $0.07/M.Which plan do I need?
Available on Go and above.
How do I switch to it?
Run
cmd --model z-ai/glm-5.3-flash, or type /model in a session and pick it. You can switch mid-session without losing context.Ship code that matches your taste
Command Code is the AI coding agent that continuously learns your taste. Start for $1.
Benchmarks from Artificial Analysis (v4.1)commandcode.ai/models/glm-5-3-flash