Qwen3-ASR Flash
Speech transcription model for accurate audio-to-text and captioning workflows
Context
53,248
Max output
4,096
Input / 1M
$0.03
Output / 1M
$0.03
Providers
2
01Profile
Profile
Released
2025-09-08
Last updated
2025-09-08
Knowledge cutoff
2024-04
Weights
Closed
02Capabilities
Capabilities
Input
audio
Output
text
Reasoning
No
Tool calling
No
Structured output
No
Attachments
No
Cache read
Not listed
Cache write
Not listed
03Pricing
Pricing across 2 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
2+ recent events since 2025-09-08
2025-10-14Addedalibabaoffering→ new
2025-10-14Addedalibaba-cnoffering→ new