GLM-5.2 Fast
TL;DR
Open-sourceHigh-throughput GLM-5.2 with a 1M context window.
Available on
Open-source models are available on every plan, including Go ($1/mo).
Switch with
Pick GLM-5.2 Fast from the selector.
Intelligence Index
Not yet scored
Speed
—
tokens / sec
Input
$3
per M tokens
Output
$10.25
per M tokens
GLM-5.2 Fast in Command Code
GLM-5.2 Fast is Z AI's open-source model — high-throughput GLM-5.2 with 1M context. It runs in Command Code with a 1M-token context window, switchable any time with /model.
GLM-5.2 Fast has no published Intelligence Index yet — it renders as "not yet scored" rather than borrowing a number.
GLM-5.2 Fast specs at a glance
What GLM-5.2 Fast accepts, how much it can hold in context, and what you need to run it.
| Spec | GLM-5.2 Fast |
|---|---|
| Context window | 1M tokens |
| Input modalities | text |
| Reasoning | No |
| Minimum plan | Go |
GLM-5.2 Fast vs the Command Code lineup
GLM-5.2 Fast alongside its nearest real alternatives in the lineup — every price straight from the billing tables.
| Model | Intelligence | Coding | Speed | Input $/M | Output $/M | Blended $/M | Context |
|---|---|---|---|---|---|---|---|
| Claude Fable 5 | 59.9 | 76.5 | ~62 tok/s | $10.00 | $50.00 | $20 | 1M |
| Grok 4.5 | 53.8 | 72.4 | ~114 tok/s | $2.00 | $6.00 | $3 | 500K |
| GLM-5.2 | 51.1 | 68.8 | ~208 tok/s | $1.40 | $4.40 | $2.15 | 1M |
| Qwen 3.6 Max Preview | 40 | — | ~46 tok/s | $1.30 | $7.80 | $2.925 | 200K |
| Kimi K3 | — | — | — | $3.00 | $15.00 | $6 | 1M |
| Kimi K2.7 Code HighSpeed | — | — | — | $1.90 | $8.00 | $3.425 | 262K |
| GLM-5.2 Fast (this page) | — | — | — | $3.00 | $10.25 | $4.8125 | 1M |
What GLM-5.2 Fast is best for
GLM-5.2 Fast is a open-source model from Z AI, priced at $4.8125 blended per million tokens.
The honest way to place it: run your own session with /model and compare against the lineup table above — the numbers on this page update as the registry and billing tables change.
When to switch away from GLM-5.2 Fast
No single model wins every task. These are GLM-5.2 Fast's computed nearest alternatives — one step up, one step down in cost, one for speed, one from the same family — each switchable mid-session with /model.
Switch to Grok 4.5
Grok 4.5 runs about 38% cheaper blended ($3 versus $4.8125 per million tokens) while scoring 53.8 on the Intelligence Index. Switch down for high-volume work where GLM-5.2 Fast's edge isn't earning its rate.
Switch to GPT-5.6 Terra
GPT-5.6 Terra streams ~155 tokens/sec at a comparable blended cost ($5.625 per million tokens). Use it when iteration speed matters more than squeezing out the last point of quality.
Switch to GLM-5.2
GLM-5.2 is the nearest Z AI sibling — same house style, lower price point ($2.15 versus $4.8125 blended per million tokens). The natural swap when you want to stay in the family.
What you pay for GLM-5.2 Fast
GLM-5.2 Fast is billed per token at the rates below — the same billing tables the Usage page charges against, so this page cannot quote a different price than you pay.
Blended cost (3:1 input:output, the shape of a typical coding session) works out to $4.8125 per million tokens.
| Per 1M tokens | Input | Output | Cache read |
|---|---|---|---|
| All requests | $3.00 | $10.25 | $0.50 |
In Command Code: caching and taste-1
Open-source models are routed across multiple upstream providers for high availability. The price you see is the mean per-provider rate; the Usage page reflects what was actually charged.
Where supported by the upstream, prompt caching is on by default — cache reads are billed at $0.50 per million tokens versus $3.00 for fresh input.
taste-1 sits between the model and the agent loop, rewriting and reranking candidate edits to match your codebase conventions.
Plan availability
GLM-5.2 Fast is an open-source model, available on every plan including Go.
Command Code is a subscription with model usage at API rates. Each plan ships with monthly LLM credits; credits roll over and never expire, and auto top-up keeps you running if you go over.
| Plan | Price/mo | LLM credits | Models |
|---|---|---|---|
| Go | $1 | $10 | Open-source only |
| Pro | $15 | $30 | Open-source + premium |
| Provider | $15 | Pay as you go | Open-source + premium |
| Max 10× | $100 | $150 | Open-source + premium |
| Max 20× | $200 | $300 | Open-source + premium |
| Teams Pro | $40 / seat | $40 / seat | Open-source + premium |
| Enterprise | Custom | Custom | Custom pool, SSO, audit logs |
Switching models with /model
In an interactive Command Code session, run /model to open the model selector. Pick GLM-5.2 Fast and it applies to this session and to future sessions until you change it again. Premium models require Pro or higher; open-source models are available on every plan, including Go.
cmd # start an interactive session
/model # open the selector and pick GLM-5.2 FastAll Command Code models, ranked by quality and speed
Quality is the Intelligence Index — an aggregate score across reasoning, math, coding, and knowledge evaluations. Speed is measured output tokens per second. Models without a published score are noted. This table is regenerated from the model registry, so it is always current.
| Model | Tier | Intelligence Index | Output speed |
|---|---|---|---|
| Claude Fable 5 | Premium | 59.9 | ~62 tok/s |
| GPT-5.6 Sol | Premium | 58.9 | ~77 tok/s |
| Claude Opus 4.8 | Premium | 55.7 | ~55 tok/s |
| GPT-5.6 Terra | Premium | 55 | ~155 tok/s |
| GPT-5.5 | Premium | 54.8 | ~81 tok/s |
| Grok 4.5 | Open-source | 53.8 | ~114 tok/s |
| Claude Opus 4.7 | Premium | 53.5 | ~52 tok/s |
| Claude Sonnet 5 | Premium | 53.4 | ~79 tok/s |
| GPT-5.4 | Premium | 51.4 | ~164 tok/s |
| GPT-5.6 Luna | Premium | 51.2 | ~234 tok/s |
| GLM-5.2 | Open-source | 51.1 | ~208 tok/s |
| Muse Spark 1.1 | Premium | 50.6 | ~130 tok/s |
| Gemini 3.5 Flash | Premium | 50.2 | ~236 tok/s |
| Claude Sonnet 4.6 | Premium | 47.2 | ~55 tok/s |
| Qwen 3.7 Max | Open-source | 46 | ~196 tok/s |
| MiniMax M3 | Open-source | 44.4 | ~113 tok/s |
| DeepSeek V4 Pro | Open-source | 44.3 | ~62 tok/s |
| GPT-5.3 Codex | Premium | 44.3 | ~106 tok/s |
| Kimi K2.6 | Open-source | 44.2 | ~43 tok/s |
| MiMo V2.5 Pro | Open-source | 42.2 | ~56 tok/s |
| Kimi K2.7 Code | Open-source | 41.9 | ~46 tok/s |
| Tencent Hy3 | Open-source | 41.2 | ~58 tok/s |
| DeepSeek V4 Flash | Open-source | 40.3 | ~106 tok/s |
| GLM-5.1 | Open-source | 40.2 | ~81 tok/s |
| Qwen 3.6 Max Preview | Open-source | 40 | ~46 tok/s |
| GPT-5.4 Mini | Premium | 40 | ~171 tok/s |
| Qwen 3.6 Plus | Open-source | 39.6 | ~53 tok/s |
| GLM-5 | Open-source | 39.5 | ~51 tok/s |
| Qwen 3.7 Plus | Open-source | 39 | ~52 tok/s |
| Kimi K2.5 | Open-source | 38.1 | ~51 tok/s |
| MiniMax M2.7 | Open-source | 38.1 | ~49 tok/s |
| Nemotron 3 Ultra | Open-source | 37.8 | ~204 tok/s |
| MiMo V2.5 | Open-source | 37.2 | ~88 tok/s |
| MiniMax M2.5 | Open-source | 33.7 | ~78 tok/s |
| Step 3.7 Flash | Open-source | 30.3 | ~407 tok/s |
| Step 3.5 Flash | Open-source | 26 | ~207 tok/s |
| Gemini 3.1 Flash Lite | Premium | 25 | ~300 tok/s |
| Claude Haiku 4.5 | Premium | 23.7 | ~103 tok/s |
| Kimi K3 | Open-source | Not yet scored | — |
| Kimi K2.7 Code HighSpeed | Open-source | Not yet scored | — |
| GLM-5.2 Fast (this page) | Open-source | Not yet scored | — |
| Inkling | Open-source | Not yet scored | — |
| Fugu Ultra | Premium | Not yet scored | — |
Frequently asked questions
GLM-5.2 Fast or Grok 4.5?
Grok 4.5 is about 38% cheaper blended ($3 vs $4.8125 per million tokens), scoring 53.8 on the Intelligence Index. Use Grok 4.5 for volume work and GLM-5.2 Fast where its edge earns the difference.
How much does GLM-5.2 Fast cost in Command Code?
$3.00 per million input tokens and $10.25 per million output tokens, with cache reads at $0.50. In an agent loop, cached context brings effective input to roughly $1.25 per million tokens.
What plan do I need for GLM-5.2 Fast?
GLM-5.2 Fast is available on every plan, including Go at $1/mo.
Does GLM-5.2 Fast support image input and reasoning?
GLM-5.2 Fast is text-only; paste code and logs rather than screenshots. It does not expose a reasoning mode.
Which Command Code model should I use?
Claude Fable 5 currently leads the lineup on the Intelligence Index (59.9). Grok 4.5 (53.8) leads the open-weights tier, available on every plan. For fast lookups, Step 3.7 Flash streams ~407 tok/s. There is no single right answer — switch per session with /model and let the task pick the model.
Can I mix GLM-5.2 Fast with other models in a workflow?
Yes. Switch per session using /model. Common pattern: keep a default model and switch up for hard problems or down for quick lookups as the task calls for it.
Are open-source model prices fixed?
Open-source models are routed across multiple upstream providers for high availability. The price listed for each is the mean per-provider rate. Actual cost on a given request may vary slightly. The Usage page reflects the price charged.
GLM-5.2 Fast or GLM-5.2?
Same model, same 1M-token context window — Fast runs on higher-throughput serving at a higher per-token rate ($3.00 vs $1.40 input). Use Fast when you are waiting on the output; use standard GLM-5.2 for unattended or cost-sensitive runs.
Does GLM-5.2 Fast handle 1M tokens?
Yes — it serves the same 1M-token context window as GLM-5.2, suited to repo-wide and long-horizon work. Command Code's auto-compact uses that window.
Is Command Code free to try?
The Go plan starts at $1/mo with $10 in LLM credits. It covers open-source models only. Pro at $15/mo unlocks premium models with $30 in LLM credits.
Does Command Code train on my code?
No. Command Code does not train on your code or store your code snippets. taste-1 data is stored locally in your project directory.
Where can I track my usage?
The Usage page in Studio shows per-request cost, token counts, and which model ran. Settings > Billing lets you change plans, buy credits, or enable auto top-up.
Does Command Code replace my editor?
No. Command Code is editor-agnostic — it runs as a CLI and works alongside any editor (Cursor, VS Code, Zed, JetBrains, Neovim, etc.).
Related reading
- GLM-5.2 in Command Codepowerful coding with 1M context and long-horizon tasks
- GLM-5.1 in Command Codelong-horizon autonomous coding agent
- GLM-5 in Command Codemulti-mode thinking & long-range planning
- Kimi K3 in Command Codelong-horizon coding & knowledge work with 1M context
- Every model, one referenceThe docs list of all Command Code models with ids and context windows.
- Pricing, limits, and dealsThe canonical price table, running deals, and usage estimates.
Ship code that matches your taste
Command Code is the AI coding agent that continuously learns your taste. Start for $1.