Gemini 3.1 Flash Lite

google/gemini-3.1-flash-lite

high-volume workhorse model with implicit caching.

cmd --model google/gemini-3.1-flash-lite
Intelligence index
15.6
Coding index
34.7
Input
$0.25 /M
Output
$1.50 /M
Cache read
$0.03 /M
Agent-loop cost
$0.10 /M in
Context window
1M tokens
Released
March 3, 2026
Modalities
→

vs. the lineup

Gemini 3.1 Flash Lite beside its stablemates and nearest rivals. The ◆ marks the best value in each column across every row shown.

pin a rival:
ModelIntelligenceCodingInput $/MOutput $/MBlended $/MContext
Gemini 3.8 Flash40.9◆76.3◆$1.50$7.50$31M
Gemini 3.7 Flash39.176.1$1.50$7.50$31.05M◆
Gemini 3.6 Flash3469.2$1.50$7.50$31M
Gemini 3.5 Flash32.670.1$1.50$9$3.381M
Gemini 3.5 Flash Lite22.249.3$0.30$2.50$0.851M
Gemini 3.1 Flash Lite ◆15.634.7$0.25$1.50$0.561M
DeepSeek V4 Pro (latest)pinned3668.8$0.66$1.98$0.991M
DeepSeek V4 Flash (latest)pinned3469.1$0.15◆$0.60◆$0.26◆1M

coding performance

The Intelligence Index and its sub-scores, ranked against every scored model in the catalog. A metric that has not been measured for Gemini 3.1 Flash Lite has been left empty.

Coding Index
34.7
#52 of 52 scored
Terminal-Bench
31.1
#52 of 52 scored
Intelligence Index
15.6
#67 of 68 scored
Long-context reasoning
74.3
reasoning across a long context
SciCode
43.4
scientific coding
GPQA Diamond
82.2
graduate-level QA

usage calculator

How far a month of credits goes on Gemini 3.1 Flash Lite.

Gemini 3.1 Flash Lite runs on Pro and up — from $20/mo.
Input tokensfresh prompt
800
Output tokensmodel reply
180
Cache read tokensre-read context
50K
cost / request $0.0020 · in $0.25 · out $1.50 · cache $0.03 per M
fresh input 10%output 14%cache reads 76%
Requests / 30 days
10K
$20 credits ÷ $0.0020 per request
~2.0K quick fixes~406 bug fixes~68 feature PRs

what real work costs

Real coding tasks priced end to end on Gemini 3.1 Flash Lite, from a quick lookup to a full-repo agent run.

One agent task
$0.03
180K in at 75% cache hit, 12K out
What you pay
$0.17 /M
all-in across every token that task touched
Sticker input
$0.25 /M
cache reads bill at $0.03 /M instead
TaskTokens in · outGemini 3.1 Flash LiteMiMo V2.6 ProClaude Haiku 4.5
Quick lookup / one-liner8K · 1K$0.0019$0.0012$0.0065
Review a 500-line PR60K · 4K$0.01$0.01$0.04
Fix a bug (agent loop)180K · 12K$0.03$0.03$0.12
Refactor a module320K · 20K$0.06$0.06$0.21
Full-repo agent run900K · 45K$0.14$0.13$0.49

frequently asked

How much does Gemini 3.1 Flash Lite cost?
$0.25/M input and $1.50/M output, cache reads $0.03/M. In an agent loop most input is cache-read, so the effective input rate is about $0.10/M.
Which plan do I need?
Available on Pro and above.
How do I switch to it?
Run cmd --model google/gemini-3.1-flash-lite, or type /model in a session and pick it. You can switch mid-session without losing context.

Ship code that matches your taste

Command Code is the AI coding agent that continuously learns your taste. Start for $1.

Benchmarks from Artificial Analysis (v4.3)commandcode.ai/models/gemini-3-1-flash-lite