Qwen3.7 Flash
Lightweight multimodal Qwen model for high-throughput text, image, and video tasks.
- Capability
- 144.6ECI · #80 of 148
- Input
- $0.03per 1M tokens
- Output
- $0.13per 1M tokens
- Context
- 1M131K max output
What it costs
- Input
- $0.03per million tokens
- Output
- $0.13per million tokens
- Cached input
- $0.003per million tokens
- Cache write
- $0.037per million tokens
- Over 32K tokens
- $0.10 / $0.40input / output per million
- Blended (3:1)
- $0.055cheaper than 96% of priced models
Typical monthly bills
| Monthly usage | Estimated cost |
|---|---|
| Side project2M input + 0.5M output tokens | $0.125 |
| Team assistant25M input + 5M output tokens | $1.40 |
| Production app250M input + 50M output tokens | $14.00 |
Official Alibaba API price, as listed on models.dev.
Independent benchmarks
#80 of 148 scored models
The shaded band is Epoch AI’s confidence range (142.4–147.6); the tick marks the median scored model.
Scores from Epoch AI, run independently of Alibaba (Qwen).
The details
- Lab
- Alibaba (Qwen)
- Released
- Jul 15, 2026
- Context window
- 1,000,000 tokens
- Max output
- 131,072 tokens
- Inputs
- Text, Images, Video
- Output
- Text
- Reasoning
- Thinking budget
- Tool calling
- Yes
- Structured output
- Yes
- Weights
- Proprietary
- API model ID
qwen3.7-flashon Alibaba- Availability
- 12 API providerslisted on models.dev
More from Alibaba (Qwen)
About Qwen3.7 Flash
How much does Qwen3.7 Flash cost?
Qwen3.7 Flash costs $0.03 input / $0.13 output per million tokens (official Alibaba API price). Cached input is $0.003 per million tokens, and writing to the cache costs $0.037. Requests over 32K tokens are billed at $0.10 input / $0.40 output. At a 3:1 input-to-output mix that is $0.055 per million tokens, cheaper than 96% of the 360 priced models we track.
What is the context window of Qwen3.7 Flash?
Qwen3.7 Flash accepts up to 1,000,000 tokens per request and can write up to 131,072 tokens in one response.
How good is Qwen3.7 Flash?
Epoch AI gives Qwen3.7 Flash a Capabilities Index score of 144.6 (likely range 142.4–147.6), ranking it #80 of 148 models Epoch has scored. Epoch AI benchmark results: GPQA Diamond 82.3%, FrontierMath Tiers 1–3 19.3%, OTIS Mock AIME 2024–2025 86.7%.
Is Qwen3.7 Flash open source?
No. Qwen3.7 Flash is proprietary; you use it through Alibaba’s API or partner platforms.
What inputs does Qwen3.7 Flash support?
Qwen3.7 Flash accepts text, images and video and replies in text. It is a reasoning model, supports tool calling and can return structured JSON output.
When was Qwen3.7 Flash released?
Alibaba (Qwen) released Qwen3.7 Flash on Jul 15, 2026.
What are the best alternatives to Qwen3.7 Flash?
The closest current models from other labs on capability, price and release date are DeepSeek V4 Flash (DeepSeek, ECI 146.1, $0.14 / $0.28), Gemma 4 26B A4B IT (Google, ECI 141.9, $0.10 / $0.38), GPT-5.4 nano (OpenAI, ECI 145.8, $0.20 / $1.25) and MiniMax-M3 (MiniMax, ECI 147.0, $0.30 / $1.20).