Qwen Flash
Efficient Qwen model for fast chat, extraction, and high-volume workloads
Context
1,000,000
Max output
250,000
Input / 1M
$0.02
Output / 1M
$0.22
Providers
7
01Profile
Profile
Released
2025-07-28
Last updated
2025-07-28
Knowledge cutoff
2024-04
Weights
Closed
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
No
Attachments
No
Cache read
Not listed
Cache write
Not listed
03Pricing
Pricing across 8 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2025-07-28
2026-09-11Addedofoxoffering→ new
2026-09-08Removed302aiofferingremoved→ new
2026-08-20Addedllmgateway-providersoffering→ new
2026-08-16Addedllmtroffering→ new
2026-08-06Addedofoxoffering→ new
2026-07-29Addedmerge-gatewayoffering→ new
2026-06-03Price upllmgatewayinput / 1M—→ $0.05
2026-06-03Price upllmgatewayoutput / 1M—→ $0.40
2026-04-18Price upllmgatewayinput / 1M$0.05→ new
2026-04-18Price upllmgatewayoutput / 1M$0.40→ new
2026-04-18ContextllmgatewayContext window1,000,000→ new
2026-04-18ContextllmgatewayMax output32,000→ new