gemini-2.0-flash-lite
Low-latency Gemini model for high-volume multimodal and agent workloads
Context
2,000,000
Max output
8,192
Input / 1M
$0.05
Output / 1M
$0.21
Providers
3
01Profile
Profile
Released
2025-06-16
Last updated
2025-08-05
Knowledge cutoff
2024-11
Weights
Closed
02Capabilities
Capabilities
Input
text · image
Output
text
Reasoning
No
Tool calling
No
Structured output
No
Attachments
Yes
Cache read
—
Cache write
—
03Pricing
Pricing across 2 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2025-06-16
2026-08-10Removedgoogleofferingremoved→ new
2026-06-24Removedllmgatewayofferingremoved→ new
2026-06-08Removedgoogle-vertexofferingremoved→ new
2026-06-08Removedvercelofferingremoved→ new
2026-06-03Price upgoogle-vertexinput / 1M—→ $0.07
2026-06-03Price upgoogle-vertexoutput / 1M—→ $0.30
2026-06-03Price upllmgatewayinput / 1M—→ $0.07
2026-06-03Price upllmgatewayoutput / 1M—→ $0.30
2026-06-03Removednano-gptofferingremoved→ new
2026-06-02DeprecatedgoogleStatus—→ deprecated
2026-05-18Addedverceloffering→ new
2026-04-18Price upllmgatewayinput / 1M$0.08→ new