GPT-3.5 Turbo 0125
Compact GPT model for low-latency assistance and high-volume workloads
Context
16,384
Max output
16,384
Input / 1M
$0.50
Output / 1M
$1.50
Providers
2
01Profile
Profile
Released
2024-01-25
Last updated
2024-01-25
Knowledge cutoff
2021-08
Weights
Closed
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
No
Tool calling
No
Structured output
No
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 2 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
2+ recent events since 2024-01-25
2026-07-25DeprecatedazureStatus—→ deprecated
2025-06-11Addedazureoffering→ new