ZMIME
Alibaba (Qwen) · Released Apr 27, 2026

Qwen3.6 Flash

Qwen vision-language model for visual reasoning, documents, and agent tasks.

  • Proprietary
  • Reasoning
  • Tool calling
  • Vision
  • Video input
Capability
143.3ECI · #85 of 148
Input
$0.188per 1M tokens
Output
$1.13per 1M tokens
Context
1M66K max output
Pricing

What it costs

Input
$0.188per million tokens
Output
$1.13per million tokens
Cache write
$0.234per million tokens
Blended (3:1)
$0.422cheaper than 65% of priced models

Typical monthly bills

Monthly usageEstimated cost
Side project2M input + 0.5M output tokens$0.938
Team assistant25M input + 5M output tokens$10.31
Production app250M input + 50M output tokens$103.13

Official Alibaba API price, as listed on models.dev.

Capability

Independent benchmarks

143.3Epoch Capabilities Index
#85 of 148 scored models
88Median 146167

The shaded band is Epoch AI’s confidence range (140.9–144.7); the tick marks the median scored model.

  • GPQA DiamondGraduate-level science questions83.3%
  • FrontierMath Tiers 1–3Research-level mathematics22.5%
  • OTIS Mock AIME 2024–2025Competition mathematics84.4%
  • SimpleQA VerifiedShort factual questions15.9%

Scores from Epoch AI, run independently of Alibaba (Qwen).

Specifications

The details

Lab
Alibaba (Qwen)
Released
Apr 27, 2026
Context window
1,000,000 tokens
Max output
65,536 tokens
Inputs
Text, Images, Video
Output
Text
Reasoning
Thinking budget
Tool calling
Yes
Structured output
Yes
Weights
Proprietary
API model ID
qwen3.6-flashon Alibaba
Availability
16 API providerslisted on models.dev

Alibaba model documentation

Same lab

More from Alibaba (Qwen)

Questions

About Qwen3.6 Flash

How much does Qwen3.6 Flash cost?

Qwen3.6 Flash costs $0.188 input / $1.13 output per million tokens (official Alibaba API price). At a 3:1 input-to-output mix that is $0.422 per million tokens, cheaper than 65% of the 360 priced models we track.

What is the context window of Qwen3.6 Flash?

Qwen3.6 Flash accepts up to 1,000,000 tokens per request and can write up to 65,536 tokens in one response.

How good is Qwen3.6 Flash?

Epoch AI gives Qwen3.6 Flash a Capabilities Index score of 143.3 (likely range 140.9–144.7), ranking it #85 of 148 models Epoch has scored. Epoch AI benchmark results: GPQA Diamond 83.3%, FrontierMath Tiers 1–3 22.5%, OTIS Mock AIME 2024–2025 84.4%, SimpleQA Verified 15.9%.

Is Qwen3.6 Flash open source?

No. Qwen3.6 Flash is proprietary; you use it through Alibaba’s API or partner platforms.

What inputs does Qwen3.6 Flash support?

Qwen3.6 Flash accepts text, images and video and replies in text. It is a reasoning model, supports tool calling and can return structured JSON output.

When was Qwen3.6 Flash released?

Alibaba (Qwen) released Qwen3.6 Flash on Apr 27, 2026.

What are the best alternatives to Qwen3.6 Flash?

The closest current models from other labs on capability, price and release date are GPT-5.4 nano (OpenAI, ECI 145.8, $0.20 / $1.25), MiniMax-M2.7 (MiniMax, ECI 145.9, $0.30 / $1.20), Gemma 4 31B IT (Google, ECI 142.8, $0.14 / $0.40) and DeepSeek V3.2 (DeepSeek, ECI 146.3, $0.296 / $0.48).