ling familyDeprecated
inclusionAI: Ling 3.0 Flash
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Context
262,144
Max output
262,144
Input / 1M
$0.02
Output / 1M
$0.06
Providers
7
01Profile
Profile
Released
2026-07-23
Last updated
2026-08-06
Knowledge cutoff
—
Weights
Closed
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
No
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 8 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2026-07-23
2026-09-14Price cutvercelinput / 1M$0.06→ $0.02
2026-09-14Price cutverceloutput / 1M$0.18→ $0.06
2026-08-20Addedllmgateway-providersoffering→ new
2026-08-20Addedllmgateway-providersoffering→ new
2026-08-12Addedllmgatewayoffering→ new
2026-08-07DeprecatedopencodeStatus—→ deprecated
2026-08-06Price cutkiloinput / 1M$0.07→ $0.06
2026-08-06Price cutkilooutput / 1M$0.22→ $0.18
2026-08-06ContextkiloContext window131,072→ 262,144
2026-08-06ContextkiloMax output16,384→ 32,768
2026-08-06Removedkiloofferingremoved→ new
2026-08-06Price upnano-gptinput / 1M$0.06→ $0.07