Googlegemini-flash-lite familyActive

gemini-2.0-flash-lite

Low-latency Gemini model for high-volume multimodal and agent workloads

JSONEmbed chartCompare
Context
2,000,000
Max output
8,192
Input / 1M
$0.05
Output / 1M
$0.21
Providers
3
01Profile

Profile

Released
2025-06-16
Last updated
2025-08-05
Knowledge cutoff
2024-11
Weights
Closed
02Capabilities

Capabilities

Input
text · image
Output
text
Reasoning
No
Tool calling
No
Structured output
No
Attachments
Yes
Cache read
Cache write
03Pricing

Pricing across 2 providers

Per 1M tokens, USD. Sorted by input price. First-party row highlighted.

ProviderProvider model idInputOutputCache readContext
poegoogle/gemini-2.0-flash-lite$0.05$0.21Not listed990,000
302aigemini-2.0-flash-lite$0.07$0.30Not listed2,000,000
qiniu-aigemini-2.0-flash-liteNot listedNot listedNot listed1,048,576
04History

Change history

12+ recent events since 2025-06-16

2026-08-10Removedgoogleofferingremoved→ new
2026-06-24Removedllmgatewayofferingremoved→ new
2026-06-08Removedgoogle-vertexofferingremoved→ new
2026-06-08Removedvercelofferingremoved→ new
2026-06-03Price upgoogle-vertexinput / 1M→ $0.07
2026-06-03Price upgoogle-vertexoutput / 1M→ $0.30
2026-06-03Price upllmgatewayinput / 1M→ $0.07
2026-06-03Price upllmgatewayoutput / 1M→ $0.30
2026-06-03Removednano-gptofferingremoved→ new
2026-06-02DeprecatedgoogleStatus→ deprecated
2026-05-18Addedverceloffering→ new
2026-04-18Price upllmgatewayinput / 1M$0.08→ new
as of 2026-09-16
About ·Corrections ·Contact ·Privacy ·EN / 中文 / 繁體
HomeModelsLabsProvidersToolsRankingsChangesNews