DeepSeek V4.1 Flash Fast
deepseek/deepseek-v4.1-flash-fast
High throughput V4.1 Flash.
cmd --model deepseek/deepseek-v4.1-flash-fast
Intelligence index
not yet scored
Coding index
not yet scored
Input
$0.16 /M
Output
$0.58 /M
Cache read
$0.02 /M
Agent-loop cost
$0.06 /M in
Context window
1M tokens
Released
September 28, 2026
Modalities
→
vs. the lineup
DeepSeek V4.1 Flash Fast beside its stablemates and nearest rivals. The ◆ marks the best value in each column across every row shown.
pin a rival:
| Model | Intelligence | Coding | Input $/M | Output $/M | Blended $/M | Context |
|---|---|---|---|---|---|---|
| DeepSeek V4.1 Flash | 39.5◆ | —◆ | $0.15◆ | $0.60◆ | $0.26◆ | 1M◆ |
| DeepSeek V4 Pro (latest) | 36◆ | 68.8◆ | $0.66◆ | $1.98◆ | $0.99◆ | 1M◆ |
| DeepSeek V4 Flash Vision (exp) | 34.8◆ | 65◆ | $0.15◆ | $0.60◆ | $0.26◆ | 1M◆ |
| DeepSeek V4 Flash (latest) | 34◆ | 69.1◆ | $0.15◆ | $0.60◆ | $0.26◆ | 1M◆ |
| DeepSeek V4.1 Flash Fast ◆ | —◆ | —◆ | $0.16◆ | $0.58◆ | $0.27◆ | 1M◆ |
| DeepSeek V4 Flash Fast | —◆ | —◆ | $0.28◆ | $0.56◆ | $0.35◆ | 1M◆ |
| Kimi K3pinned | 43.6◆ | 76.2◆ | $3◆ | $15◆ | $6◆ | 1M◆ |
| Kimi K2.7 Codepinned | 25.8◆ | 60.8◆ | $0.95◆ | $4◆ | $1.71◆ | 256K◆ |
peak hours
DeepSeek V4.1 Flash Fast is not billed at one flat rate — the price changes with the time of day. Every other price on this page is the off-peak rate, which is what a request costs 17 hours out of 24. Peak runs Monday to Friday only, so it is never charged at the weekend.
| Band | Hours (UTC) | Input $/M | Output $/M | Cache read $/M |
|---|---|---|---|---|
| Off-peak | 17h/day | $0.16 | $0.58 | $0.02 |
| Peak | 7h/day · 01–04 & 06–10 UTC, Mon–Fri | $0.32 | $1.16 | $0.03 |
usage calculator
How far a month of credits goes on DeepSeek V4.1 Flash Fast.
Input tokensfresh prompt
800Output tokensmodel reply
180Cache read tokensre-read context
50Kcost / request $0.0010 · in $0.16 · out $0.58 · cache $0.02 per M
fresh input 12%output 10%cache reads 77%
Requests / 30 days
9.7K
$10 credits ÷ $0.0010 per request
~1.9K quick fixes~387 bug fixes~65 feature PRs
what real work costs
Real coding tasks priced end to end on DeepSeek V4.1 Flash Fast, from a quick lookup to a full-repo agent run.
One agent task
$0.02
180K in at 75% cache hit, 12K out
What you pay
$0.08 /M
all-in across every token that task touched
Sticker input
$0.16 /M
cache reads bill at $0.02 /M instead
| Task | Tokens in · out | DeepSeek V4.1 Flash Fast | Muse Spark 1.3 Contributor | Claude Haiku 4.5 |
|---|---|---|---|---|
| Quick lookup / one-liner | 8K · 1K | $0.0008 | $0.0003 | $0.0065 |
| Review a 500-line PR | 60K · 4K | $0.0059 | $0.0027 | $0.04 |
| Fix a bug (agent loop) | 180K · 12K | $0.02 | $0.0072 | $0.12 |
| Refactor a module | 320K · 20K | $0.03 | $0.01 | $0.21 |
| Full-repo agent run | 900K · 45K | $0.07 | $0.03 | $0.49 |
frequently asked
How much does DeepSeek V4.1 Flash Fast cost?
$0.16/M input and $0.58/M output, cache reads $0.02/M. In an agent loop most input is cache-read, so the effective input rate is about $0.06/M.Which plan do I need?
Available on Go and above.
How do I switch to it?
Run
cmd --model deepseek/deepseek-v4.1-flash-fast, or type /model in a session and pick it. You can switch mid-session without losing context.Ship code that matches your taste
Command Code is the AI coding agent that continuously learns your taste. Start for $1.
Benchmarks from Artificial Analysis (v4.3)commandcode.ai/models/deepseek-v4-1-flash-fast