OpenAI: GPT-4o-mini (2024-07-18)
Compact GPT model for low-latency assistance and high-volume workloads
Context
128,000
Max output
16,384
Input / 1M
$0.15
Output / 1M
$0.60
Providers
3
01Profile
Profile
Released
2024-07-18
Last updated
2026-06-11
Knowledge cutoff
—
Weights
Closed
02Capabilities
Capabilities
Input
text · image · pdf
Output
text
Reasoning
No
Tool calling
Yes
Structured output
Yes
Attachments
Yes
Cache read
—
Cache write
—
03Pricing
Pricing across 3 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
5+ recent events since 2024-07-18
2026-06-10ContextveniceContext window128,000→ new
2026-06-10ContextveniceMax output16,384→ new
2026-05-15Addedopenrouteroffering→ new
2026-03-15Addedkilooffering→ new
2026-03-06Addedveniceoffering→ new
05Usage
Monthly token share
Rank #62 · 2025-09
0.1%