MiMo V2.5

-98%
xiaomi/mimo-v2.5

efficient long-context agentic coding.

cmd --model xiaomi/mimo-v2.5
Intelligence index
38
Output speed
77.3 tok/s
Input
$0.14 /M
Output
$0.28 /M
Cache read
$0.0028 /M
Agent-loop cost
$0.04 /M in
Context window
1M tokens
Released
April 22, 2026
Modalities

deal spotlight

98% off, applied at billing rather than at checkout. Every price and calculator on this page already quotes the discounted rate.

−98%
permanent
Full terms →
Input $/M
$0.14
was $0.80
Output $/M
$0.28
was $4.00
Cache read $/M
$0.0028
was $0.16
No code, no toggle. Pick MiMo V2.5 from /model in the CLI and the discounted rate applies automatically.

vs. the lineup

MiMo V2.5 beside its stablemates and nearest rivals. The ◆ marks the best value in each column across every row shown.

pin a rival:
ModelIntelligenceCodingSpeedInput $/MOutput $/MBlended $/MContext
Muse Spark 1.2 Contributor56.872.2$0.10$0.20$0.131.05M
MiMo V2.5 Pro42.960.251.4$0.43$0.87$0.541M
MiniMax M2.738.952.6$0.30$1.20$0.52200K
Nemotron 3 Ultra38.349.3122.2$0.60$2.40$1.051M
MiMo V2.53856.877.3$0.14$0.28$0.181M
Step 3.7 Flash30.939.6391.3$0.20$1.15$0.44256K
DeepSeek V4 Pro (latest)pinned45.359.463.3$0.66$1.98$0.991M
DeepSeek V4 Flash (latest)pinned5269.1115.9$0.22$0.66$0.331M

coding performance

The Intelligence Index and its sub-scores, ranked against every scored model in the catalog. A metric that has not been measured for MiMo V2.5 has been left empty.

Coding Index
56.8
#28 of 40 scored
Terminal-Bench
63.7
#28 of 40 scored
Intelligence Index
38
#39 of 46 scored
Long-context reasoning
68.3
reasoning across a long context
SciCode
43.1
scientific coding
GPQA Diamond
84.9
graduate-level QA

usage calculator

How far a month of credits goes on MiMo V2.5.

Input tokensfresh prompt
800
Output tokensmodel reply
180
Cache read tokensre-read context
50K
cost / request $0.0003 · in $0.14 · out $0.28 · cache $0.0028 per M
fresh input 37%output 17%cache reads 46%
Requests / 30 days
33K
$10 credits ÷ $0.0003 per request
~6.6K quick fixes~1.3K bug fixes~220 feature PRs

what real work costs

Real coding tasks priced end to end on MiMo V2.5, from a quick lookup to a full-repo agent run.

One agent task
$0.01
180K in at 75% cache hit, 12K out
What you pay
$0.05 /M
all-in across every token that task touched
Sticker input
$0.14 /M
cache reads bill at $0.0028 /M instead
TaskTokens in · outMiMo V2.5Muse Spark 1.2 ContributorClaude Haiku 4.5
Quick lookup / one-liner8K · 1K$0.0004$0.0003$0.0065
Review a 500-line PR60K · 4K$0.0038$0.0027$0.04
Fix a bug (agent loop)180K · 12K$0.01$0.0072$0.12
Refactor a module320K · 20K$0.02$0.01$0.21
Full-repo agent run900K · 45K$0.04$0.03$0.49

frequently asked

What is MiMo V2.5 best for?
MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding tasks. Its 1M context window supports complete documents, extended conversations, and complex task contexts in a single pass, making it ideal for integration with agent frameworks where strong reasoning, rich perception, and cost efficiency all matter.
How much does MiMo V2.5 cost?
$0.14/M input and $0.28/M output, cache reads $0.0028/M. In an agent loop most input is cache-read, so the effective input rate is about $0.04/M.
Which plan do I need?
Available on Go and above.
How do I switch to it?
Run cmd --model xiaomi/mimo-v2.5, or type /model in a session and pick it. You can switch mid-session without losing context.

Ship code that matches your taste

Command Code is the AI coding agent that continuously learns your taste. Start for $1.

Benchmarks from Artificial Analysis (v4.1)commandcode.ai/models/mimo-v2-5