gemini-2.5-flash-lite-preview-06-17
Low-latency Gemini model for high-volume multimodal and agent workloads
Context
1,048,576
Max output
65,536
Input / 1M
$0.09
Output / 1M
$0.36
Providers
2
01Profile
Profile
Released
2026-01
Last updated
2026-01
Knowledge cutoff
—
Weights
Closed
02Capabilities
Capabilities
Input
text · video · image · audio
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
Yes
Attachments
Yes
Cache read
—
Cache write
—
03Pricing
Pricing across 2 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
10+ recent events since 2026-01
2026-09-05Contextnano-gptContext window1,048,756→ 1,048,576
2026-06-08Removedgoogle-vertexofferingremoved→ new
2026-05-19Removedgoogleofferingremoved→ new
2026-02-17Addednano-gptoffering→ new
2026-01-30Addedjiekouoffering→ new
2025-12-10ContextgoogleContext window65,536→ 1,048,576
2025-09-27ContextgoogleContext window1,048,576→ 65,536
2025-09-23ContextgoogleContext window65,536→ 1,048,576
2025-06-23Addedgoogle-vertexoffering→ new
2025-06-18Addedgoogleoffering→ new