GoogleActive
gemini-2.5-flash-lite-preview-09-2025
Low-latency Gemini model for high-volume multimodal and agent workloads
Context
1,048,576
Max output
65,536
Input / 1M
$0.09
Output / 1M
$0.36
Providers
3
01Profile
Profile
Released
2025-09-26
Last updated
2026-01
Knowledge cutoff
2025-01
Weights
Closed
02Capabilities
Capabilities
Input
text · image
Output
text
Reasoning
No
Tool calling
Yes
Structured output
No
Attachments
Yes
Cache read
—
Cache write
—
03Pricing
Pricing across 4 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2025-09-26
2026-09-05Contextnano-gptContext window1,048,756→ 1,048,576
2026-09-05Contextnano-gptContext window1,048,756→ 1,048,576
2026-08-02Removedkiloofferingremoved→ new
2026-07-09Removedllmgatewayofferingremoved→ new
2026-07-09Removedopenrouterofferingremoved→ new
2026-06-24Addedllmgatewayoffering→ new
2026-06-08Removedvercelofferingremoved→ new
2026-05-19Removedgoogleofferingremoved→ new
2026-05-19Removedgoogle-vertexofferingremoved→ new
2026-05-19Removedllmgatewayofferingremoved→ new
2026-05-15ContextopenrouterMax output65,536→ 65,535
2026-04-18Price upllmgatewayinput / 1M$0.10→ new