Gemini 3.1 Flash Lite
Low-latency Gemini model for high-volume multimodal and agent workloads
Context
1,048,576
Max output
65,536
Input / 1M
$0.13
Output / 1M
$0.75
Providers
25
01Profile
Profile
Released
2026-05-07
Last updated
2026-05-07
Knowledge cutoff
2025-01
Weights
Closed
02Capabilities
Capabilities
Input
text · image · video · audio · pdf
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
Yes
Attachments
Yes
Cache read
$0.03
Cache write
Not listed
03Pricing
Pricing across 29 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2026-05-07
2026-09-08Added302aioffering→ new
2026-09-02ContextcortecsMax output1,048,576→ 65,535
2026-08-31Price cutkiloinput / 1M$0.25→ $0.13
2026-08-31Price cutkilooutput / 1M$1.50→ $0.75
2026-08-27Addededenaioffering→ new
2026-08-27Addededenaioffering→ new
2026-08-27Addededenaioffering→ new
2026-08-27Addedorcarouteroffering→ new
2026-08-27Price uprequestyinput / 1M$0.23→ $0.25
2026-08-27Price uprequestyoutput / 1M$1.35→ $1.50
2026-08-27Price uprequestyinput / 1M$0.25→ $0.28
2026-08-27Price uprequestyoutput / 1M$1.49→ $1.65