Qwen 3.7 Flash
qwen/qwen3.7-flash
fast low-cost agentic coding & reasoning.
cmd --model qwen/qwen3.7-flash
Intelligence index
not yet scored
Output speed
not yet scored
Input
$0.03 /M
Output
$0.13 /M
Cache read
$0.006 /M
Agent-loop cost
$0.01 /M in
Context window
1M tokens
Released
—
Modalities
→
vs. the lineup
Qwen 3.7 Flash beside its stablemates and nearest rivals. The ◆ marks the best value in each column across every row shown.
pin a rival:
| Model | Intelligence | Coding | Speed | Input $/M | Output $/M | Blended $/M | Context |
|---|---|---|---|---|---|---|---|
| Qwen 3.7 Max | 46.7◆ | 66◆ | —◆ | $2.50◆ | $7.50◆ | $3.75◆ | 1M◆ |
| Qwen 3.6 Max Preview | 41.1◆ | —◆ | —◆ | $1.30◆ | $7.80◆ | $2.92◆ | 200K◆ |
| Qwen 3.6 Plus | 40.5◆ | 54.5◆ | —◆ | $0.50◆ | $3◆ | $1.13◆ | 200K◆ |
| Qwen 3.7 Plus | 39.4◆ | 55.9◆ | 54.9◆ | $0.40◆ | $1.60◆ | $0.70◆ | 1M◆ |
| Qwen 3.7 Flash ◆ | —◆ | —◆ | —◆ | $0.03◆ | $0.13◆ | $0.06◆ | 1M◆ |
| Qwen 3.8 Max | —◆ | —◆ | —◆ | $2◆ | $6◆ | $3◆ | 1M◆ |
| DeepSeek V4 Pro (latest)pinned | 45.3◆ | 59.4◆ | 63.3◆ | $0.43◆ | $0.87◆ | $0.54◆ | 1M◆ |
| DeepSeek V4 Flash (latest)pinned | 52◆ | 69.1◆ | 115.9◆ | $0.14◆ | $0.28◆ | $0.18◆ | 1M◆ |
price bands
Qwen 3.7 Flash is not billed at one flat rate — the price changes with how much context a request carries. The header quotes the standard band; here is every band.
| Band | Context | Input $/M | Output $/M | Cache read $/M |
|---|---|---|---|---|
| Standard | ≤ 32K | $0.03 | $0.13 | $0.006 |
| Extended 1 | ≤ 256K | $0.10 | $0.40 | $0.02 |
| Long context | > 256K | $0.20 | $0.80 | $0.04 |
usage calculator
How far a month of credits goes on Qwen 3.7 Flash.
Input tokensfresh prompt
800Output tokensmodel reply
180Cache read tokensre-read context
50Kcost / request $0.0003 · in $0.03 · out $0.13 · cache $0.006 per M
fresh input 7%output 7%cache reads 86%
Requests / 30 days
29K
$10 credits ÷ $0.0003 per request
~5.8K quick fixes~1.2K bug fixes~192 feature PRs
what real work costs
Real coding tasks priced end to end on Qwen 3.7 Flash, from a quick lookup to a full-repo agent run.
One agent task
$0.0037
180K in at 75% cache hit, 12K out
What you pay
$0.02 /M
all-in across every token that task touched
Sticker input
$0.03 /M
cache reads bill at $0.006 /M instead
| Task | Tokens in · out | Qwen 3.7 Flash | Claude Haiku 4.5 |
|---|---|---|---|
| Quick lookup / one-liner | 8K · 1K | $0.0002 | $0.0065 |
| Review a 500-line PR | 60K · 4K | $0.0013 | $0.04 |
| Fix a bug (agent loop) | 180K · 12K | $0.0037 | $0.12 |
| Refactor a module | 320K · 20K | $0.0067 | $0.21 |
| Full-repo agent run | 900K · 45K | $0.02 | $0.49 |
frequently asked
What is Qwen 3.7 Flash best for?+
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world visual perception.
How much does Qwen 3.7 Flash cost?+
$0.03/M input and $0.13/M output, cache reads $0.006/M. In an agent loop most input is cache-read, so the effective input rate is about $0.01/M.Which plan do I need?+
Available on Go and above.
How do I switch to it?+
Run
cmd --model qwen/qwen3.7-flash, or type /model in a session and pick it. You can switch mid-session without losing context.Ship code that matches your taste
Command Code is the AI coding agent that continuously learns your taste. Start for $1.
Benchmarks from Artificial Analysis (v4.1)commandcode.ai/models/qwen3-7-flash