Qwen3.8 Flash Next
Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding.
- Capability
- —No ECI score yet
- Input
- $0.20per 1M tokens
- Output
- $0.50per 1M tokens
- Context
- 262K131K max output
What it costs
- Input
- $0.20per million tokens
- Output
- $0.50per million tokens
- Blended (3:1)
- $0.275cheaper than 73% of priced models
Typical monthly bills
| Monthly usage | Estimated cost |
|---|---|
| Side project2M input + 0.5M output tokens | $0.65 |
| Team assistant25M input + 5M output tokens | $7.50 |
| Production app250M input + 50M output tokens | $75.00 |
Median of 5 paid API providers on models.dev.
Independent benchmarks
Epoch AI has not published a Capabilities Index score for Qwen3.8 Flash Next yet.
The details
- Lab
- Alibaba (Qwen)
- Released
- Aug 27, 2026
- Context window
- 262,144 tokens
- Max output
- 131,072 tokens
- Inputs
- Text, Images, Video
- Output
- Text
- Reasoning
- Always on
- Tool calling
- Yes
- Structured output
- Yes
- Availability
- 5 API providerslisted on models.dev
What Alibaba (Qwen) claims
Published by the lab at launch. Settings vary, so compare these only with care.
| Benchmark | Score | Setting | Source |
|---|---|---|---|
| DeepSWE v1.1 | 58.7 | — | Source |
| SWE-Bench Pro | 62.5 | — | Source |
| SWE-Bench Multilingual | 81 | — | Source |
| NL2Repo | 48.1 | — | Source |
| CoWorkBench | 73.9 | — | Source |
| JobBench | 55.7 | — | Source |
| Agents' Last Exam | 24.3 | — | Source |
| Agents' Last Exam | 51.2 | — | Source |
| Toolathlon-Verified | 73.5 | — | Source |
| IFBench | 81.3 | — | Source |
| GPQA Diamond | 91.7 | — | Source |
| Humanity's Last Exam | 35.9 | without tools; GPT-4o judge | Source |
More from Alibaba (Qwen)
About Qwen3.8 Flash Next
How much does Qwen3.8 Flash Next cost?
Qwen3.8 Flash Next costs $0.20 input / $0.50 output per million tokens (median across 5 API providers). At a 3:1 input-to-output mix that is $0.275 per million tokens, cheaper than 73% of the 360 priced models we track.
What is the context window of Qwen3.8 Flash Next?
Qwen3.8 Flash Next accepts up to 262,144 tokens per request and can write up to 131,072 tokens in one response.
How good is Qwen3.8 Flash Next?
Epoch AI has not published a Capabilities Index score for Qwen3.8 Flash Next yet.
Is Qwen3.8 Flash Next open source?
Yes. Alibaba (Qwen) publishes the weights under the qwen-community-1.0 licence, so it can be downloaded and self-hosted.
What inputs does Qwen3.8 Flash Next support?
Qwen3.8 Flash Next accepts text, images and video and replies in text. It is a reasoning model, supports tool calling and can return structured JSON output.
When was Qwen3.8 Flash Next released?
Alibaba (Qwen) released Qwen3.8 Flash Next on Aug 27, 2026.
What are the best alternatives to Qwen3.8 Flash Next?
The closest current models from other labs on capability, price and release date are MiniCPM5-2B (OpenBMB, $0.124 / $0.743), DeepSeek V4 Flash Vision Exp (DeepSeek, $0.216 / $0.647), Hy3 (Tencent, $0.132 / $0.53) and Viv Fast (Vivgrid, $0.13 / $0.40).