Gemini 2.5 Flash Preview
Compact GPT model for low-latency assistance and high-volume workloads
Context
1,048,576
Max output
65,536
Input / 1M
$0.15
Output / 1M
$0.60
Providers
1
01Profile
Profile
Released
2025-04-17
Last updated
2025-04-17
Knowledge cutoff
—
Weights
Closed
02Capabilities
Capabilities
Input
text · image · audio
Output
text
Reasoning
Yes
Tool calling
No
Structured output
No
Attachments
Yes
Cache read
—
Cache write
—
03Pricing
Pricing across 2 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2025-04-17
2026-09-05Contextnano-gptContext window1,048,756→ 1,048,576
2026-09-05Contextnano-gptContext window1,048,756→ 1,048,576
2026-05-19Removedgoogleofferingremoved→ new
2026-05-19Removedgoogle-vertexofferingremoved→ new
2026-04-17Price upgoogle-vertexinput / 1M$0.15→ new
2026-04-17Price upgoogle-vertexoutput / 1M$0.60→ new
2026-04-17Contextgoogle-vertexContext window1,048,576→ new
2026-04-17Contextgoogle-vertexMax output65,536→ new
2026-02-17Addednano-gptoffering→ new
2026-02-17Addednano-gptoffering→ new
2025-07-16Removedopenrouterofferingremoved→ new
2025-06-24Price upopenrouterinput / 1M—→ $0.15