Alibabaqwen familyActive

Qwen3.8 Flash Next

Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding

JSONEmbed chartCompare
Context
262,144
Max output
262,144
Input / 1M
$0.15
Output / 1M
$0.47
Providers
3
01Profile

Profile

Released
2026-08-27
Last updated
2026-08-27
Knowledge cutoff
Weights
Open
02Capabilities

Capabilities

Input
text · image
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
Yes
Attachments
Yes
Cache read
Cache write
03Pricing

Pricing across 4 providers

Per 1M tokens, USD. Sorted by input price. First-party row highlighted.

ProviderProvider model idInputOutputCache readContext
amdQwen3.8-Flash-Next$0.15$0.47$0.02262,144
requestyqwen3.8-flash-next$0.20$0.50$0.05262,144
requestyqwen3.8-flash-next@eu@eu$0.20$0.50$0.05262,144
cortecsqwen3.8-flash-next$0.20$0.50$0.05262,144
04History

Change history

6+ recent events since 2026-08-27

2026-09-14Removedvercelofferingremoved→ new
2026-09-02Addedrequestyoffering→ new
2026-09-02Addedrequestyoffering→ new
2026-08-31Addedverceloffering→ new
2026-08-28Addedamdoffering→ new
2026-08-28Addedcortecsoffering→ new
as of 2026-09-16
About ·Corrections ·Contact ·Privacy ·EN / 中文 / 繁體
HomeModelsLabsProvidersToolsRankingsChangesNews