Model

Gemini 3.1 Flash-Lite NO.19

Cheapest Gemini, optimized for low-latency high-volume traffic.

Google DeepMind · Gemini · Fast · Closed weights

Specification

Fig.01Gemini — Google DeepMind

USA

Context window
1.0Mtokens
Max output
66Ktokens
Input price
$0.25/ 1M
Output price
$1.50/ 1M
Cached input
$0.05/ 1M
Throughput
Class
Fast
Modalities
Text, Image, Audio, Video, Pdf
Weights
Closed
Knowledge cutoff
Jan 2025
Released
May 2026

Overview

Gemini 3.1 Flash-Lite is Google's most cost-efficient Gemini tier, tuned for ultra-low-latency, high-volume workloads. It keeps a full 1M-token multimodal context and posts surprisingly strong scores for its class, but it is optimized for cost and speed, not reasoning depth. It is the default choice when throughput, latency, and cost matter more than intelligence.

Strengths

  • Extremely low price per token
  • Very high throughput and low latency
  • Full 1M-token multimodal context
  • Solid quality for simple, high-volume tasks

Trade-offs

  • Weak on hard reasoning and complex coding
  • Not suited to agentic autonomy
  • Trails Flash and Pro on complex tasks

Fit

Best for

  • High-volume classification and extraction
  • Cheap summarization and chat
  • Latency-sensitive consumer features
  • Routing and pre-processing pipelines
  • Bulk multimodal tagging

Not ideal for

  • Frontier reasoning or research
  • Any autonomous agent workflow

Capabilities

Ten axes, normalized 0–100 and scored the same way across the whole catalog — so a 78 here means what a 78 means anywhere else on the bench.

Fig.02Capability radar

0 — 100

Reasoning
70
Coding
62
Math
72
Writing
75
Knowledge
78
Speed
96
Agentic
55
Vision
82
Multilingual
82
Long Context
86

Benchmarks

Public results, with independent third-party runs marked. Bars normalize percentages against 100 and Elo ratings against a 1500 ceiling.

Independent index

Artificial Analysis Intelligence Index

Composite of ~9–10 independent evals · Artificial Analysis

25/ 100

BenchmarkResult
Artificial Analysis Intelligence IndexIndependentGeneral · Artificial Analysis
25%
GPQA DiamondIndependentReasoning · Independent aggregators
86.9%
MMLU-ProKnowledge · Vendor-reported
83%
MMMUVision · Vendor-reported
76.8%
LMArena EloIndependentGeneral · LMArena
1432Elo

Alternatives

Models in roughly the same class — the ones worth weighing against this record.

Index

All models/EDUCATION