Model
Gemini 3.1 Flash-Lite NO.19
Cheapest Gemini, optimized for low-latency high-volume traffic.
Google DeepMind · Gemini · Fast · Closed weights
Specification
11 fields
Fig.01 — Gemini — Google DeepMind
USA
- Context window
- 1.0Mtokens
- Max output
- 66Ktokens
- Input price
- $0.25/ 1M
- Output price
- $1.50/ 1M
- Cached input
- $0.05/ 1M
- Throughput
- —
- Class
- Fast
- Modalities
- Text, Image, Audio, Video, Pdf
- Weights
- Closed
- Knowledge cutoff
- Jan 2025
- Released
- May 2026
Overview
Gemini 3.1 Flash-Lite is Google's most cost-efficient Gemini tier, tuned for ultra-low-latency, high-volume workloads. It keeps a full 1M-token multimodal context and posts surprisingly strong scores for its class, but it is optimized for cost and speed, not reasoning depth. It is the default choice when throughput, latency, and cost matter more than intelligence.
Strengths
- Extremely low price per token
- Very high throughput and low latency
- Full 1M-token multimodal context
- Solid quality for simple, high-volume tasks
Trade-offs
- Weak on hard reasoning and complex coding
- Not suited to agentic autonomy
- Trails Flash and Pro on complex tasks
Fit
Best for
- High-volume classification and extraction
- Cheap summarization and chat
- Latency-sensitive consumer features
- Routing and pre-processing pipelines
- Bulk multimodal tagging
Not ideal for
- Frontier reasoning or research
- Any autonomous agent workflow
Capabilities
Normalized 0—100
Ten axes, normalized 0–100 and scored the same way across the whole catalog — so a 78 here means what a 78 means anywhere else on the bench.
Fig.02 — Capability radar
0 — 100
- Reasoning
- 70
- Coding
- 62
- Math
- 72
- Writing
- 75
- Knowledge
- 78
- Speed
- 96
- Agentic
- 55
- Vision
- 82
- Multilingual
- 82
- Long Context
- 86
Benchmarks
05 results
Public results, with independent third-party runs marked. Bars normalize percentages against 100 and Elo ratings against a 1500 ceiling.
Independent index
Artificial Analysis Intelligence Index
Composite of ~9–10 independent evals · Artificial Analysis
25/ 100
- Artificial Analysis Intelligence IndexIndependentGeneral · Artificial Analysis
- 25%
- GPQA DiamondIndependentReasoning · Independent aggregators
- 86.9%
- MMLU-ProKnowledge · Vendor-reported
- 83%
- MMMUVision · Vendor-reported
- 76.8%
- LMArena EloIndependentGeneral · LMArena
- 1432Elo
Alternatives
02 comparable
Models in roughly the same class — the ones worth weighing against this record.
Index
Record 19 of 41