Skip to content

Gemini 3.1 Flash-Lite

Google · Released 7 May 2026 · gemini-3.1-flash-lite

Compare this model
ProprietaryBudget

$0.25

Input price / 1M tokens

#8 of 61, cheapest first

$1.50

Output price / 1M tokens

#11 of 61, cheapest first

1.05M

Context window

#11 of 71

280

Output tokens / second

#4 of 61

16%

Humanity’s Last Exam (no tools)

#18 of 23

86.9%

GPQA Diamond

#18 of 27

76.8%

MMMU-Pro

#5 of 12

72%

LiveCodeBench

#4 of 7

Summary

Gemini 3.1 Flash-Lite is a proprietary model from Google, released on 7 May 2026.

At $0.25 input and $1.50 output per million tokens, it is #8 of 61 on input price, cheapest first.

Its 1.05M-token context window ranks #11 of 71.

Measured output speed is 280 tokens per second, #4 of 61.

Low-latency, cost-efficient multimodal model for high-frequency lightweight tasks, simple data extraction and latency-sensitive apps.

Benchmarks · 8 reported

BenchmarkScoreBarRankSource
Humanity’s Last Exam (no tools)no tools, full set text + multimodal, high thinking16%#18 of 23Google Gemini 3.1 Flash-Lite evaluation report (preview model)Self-reported
GPQA Diamondno tools, high thinking86.9%#18 of 27Google Gemini 3.1 Flash-Lite evaluation report (preview model)Self-reported
MMMU-Prono tools76.8%#5 of 12Google Gemini 3.1 Flash-Lite evaluation report (preview model)Self-reported
Video-MMMU84.8%–Google Gemini 3.1 Flash-Lite evaluation report (preview model)Self-reported
SimpleQA Verified43.3%–Google Gemini 3.1 Flash-Lite evaluation report (preview model)Self-reported
MMMLU88.9%–Google Gemini 3.1 Flash-Lite evaluation report (preview model)Self-reported
LiveCodeBenchcode generation, 175 UI problems dated 1 Jan 2025 to 1 May 202572%#4 of 7Google Gemini 3.1 Flash-Lite evaluation report (preview model)Self-reported
MRCR v2 (8-needle)128k average (12.3% at 1M pointwise)60.1%#9 of 13Google Gemini 3.1 Flash-Lite evaluation report (preview model)Self-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with Gemini 3.1 Flash-Lite, and the closest scores.

Compare 4 side by side

Humanity’s Last Exam (no tools)

#18 of 23

Expert questions, no search or code

GPQA Diamond

#18 of 27

Graduate-level science questions

MMMU-Pro

#5 of 12

Hard questions that need an image

LiveCodeBench

#4 of 7

Fresh competitive programming problems

Compare Gemini 3.1 Flash-Lite side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on Gemini 3.1 Flash-Lite?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.