Qwen Flash
Efficient Qwen model for fast chat, extraction, and high-volume workloads.
- Capability
- —No ECI score yet
- Input
- $0.05per 1M tokens
- Output
- $0.40per 1M tokens
- Context
- 1M33K max output
What it costs
- Input
- $0.05per million tokens
- Output
- $0.40per million tokens
- Blended (3:1)
- $0.138cheaper than 87% of priced models
Typical monthly bills
| Monthly usage | Estimated cost |
|---|---|
| Side project2M input + 0.5M output tokens | $0.30 |
| Team assistant25M input + 5M output tokens | $3.25 |
| Production app250M input + 50M output tokens | $32.50 |
Official Alibaba API price, as listed on models.dev.
Independent benchmarks
Epoch AI has not published a Capabilities Index score for Qwen Flash yet.
The details
- Lab
- Alibaba (Qwen)
- Released
- Jul 28, 2025
- Knowledge cutoff
- Apr 2024
- Context window
- 1,000,000 tokens
- Max output
- 32,768 tokens
- Inputs
- Text
- Output
- Text
- Reasoning
- Thinking budget
- Tool calling
- Yes
- Structured output
- No
- Weights
- Proprietary
- API model ID
qwen-flashon Alibaba- Availability
- 6 API providerslisted on models.dev
More from Alibaba (Qwen)
About Qwen Flash
How much does Qwen Flash cost?
Qwen Flash costs $0.05 input / $0.40 output per million tokens (official Alibaba API price). At a 3:1 input-to-output mix that is $0.138 per million tokens, cheaper than 87% of the 360 priced models we track.
What is the context window of Qwen Flash?
Qwen Flash accepts up to 1,000,000 tokens per request and can write up to 32,768 tokens in one response.
How good is Qwen Flash?
Epoch AI has not published a Capabilities Index score for Qwen Flash yet.
Is Qwen Flash open source?
No. Qwen Flash is proprietary; you use it through Alibaba’s API or partner platforms.
What inputs does Qwen Flash support?
Qwen Flash accepts text and replies in text. It is a reasoning model, supports tool calling.
When was Qwen Flash released?
Alibaba (Qwen) released Qwen Flash on Jul 28, 2025. Its training data runs to Apr 2024.
What are the best alternatives to Qwen Flash?
The closest current models from other labs on capability, price and release date are Voxtral Small 24B 2507 (Mistral AI, $0.10 / $0.30), Apertus 8B (Swiss AI, $0.10 / $0.20), Granite-4.0-H-Small (IBM, $0.064 / $0.265) and Sarvam 105B (Sarvam AI, $0.047 / $0.186).