ZMIME
Alibaba (Qwen) · Released Sep 17, 2026

Qwen3.8 Omni Flash

Qwen omni model for text, vision, audio, and multimodal agent tasks.

  • Proprietary
  • Reasoning
  • Tool calling
  • Vision
  • Audio input
  • Video input
Capability
—No ECI score yet
Input
$0.15per 1M tokens
Output
$0.47per 1M tokens
Context
1M131K max output
Pricing

What it costs

Input
$0.15per million tokens
Output
$0.47per million tokens
Cached input
$0.016per million tokens
Blended (3:1)
$0.23cheaper than 77% of priced models

Typical monthly bills

Monthly usageEstimated cost
Side project2M input + 0.5M output tokens$0.535
Team assistant25M input + 5M output tokens$6.10
Production app250M input + 50M output tokens$61.00

Official Alibaba API price, as listed on models.dev.

Capability

Independent benchmarks

Epoch AI has not published a Capabilities Index score for Qwen3.8 Omni Flash yet.

Specifications

The details

Lab
Alibaba (Qwen)
Released
Sep 17, 2026
Context window
1,000,000 tokens
Max output
131,072 tokens
Inputs
Text, Images, Audio, Video
Output
Text
Reasoning
Adjustable effort low · medium · xhigh
Tool calling
Yes
Structured output
Yes
Weights
Proprietary
API model ID
qwen3.8-omni-flashon Alibaba
Availability
10 API providerslisted on models.dev

Alibaba model documentation

Same lab

More from Alibaba (Qwen)

Questions

About Qwen3.8 Omni Flash

How much does Qwen3.8 Omni Flash cost?

Qwen3.8 Omni Flash costs $0.15 input / $0.47 output per million tokens (official Alibaba API price). Cached input is $0.016 per million tokens. At a 3:1 input-to-output mix that is $0.23 per million tokens, cheaper than 77% of the 360 priced models we track.

What is the context window of Qwen3.8 Omni Flash?

Qwen3.8 Omni Flash accepts up to 1,000,000 tokens per request and can write up to 131,072 tokens in one response.

How good is Qwen3.8 Omni Flash?

Epoch AI has not published a Capabilities Index score for Qwen3.8 Omni Flash yet.

Is Qwen3.8 Omni Flash open source?

No. Qwen3.8 Omni Flash is proprietary; you use it through Alibaba’s API or partner platforms.

What inputs does Qwen3.8 Omni Flash support?

Qwen3.8 Omni Flash accepts text, images, audio and video and replies in text. It is a reasoning model with low, medium and xhigh effort settings, supports tool calling and can return structured JSON output.

When was Qwen3.8 Omni Flash released?

Alibaba (Qwen) released Qwen3.8 Omni Flash on Sep 17, 2026.

What are the best alternatives to Qwen3.8 Omni Flash?

The closest current models from other labs on capability, price and release date are Viv Fast (Vivgrid, $0.13 / $0.40), MiniCPM5-2B (OpenBMB, $0.124 / $0.743), Hy3 (Tencent, $0.132 / $0.53) and MiMo-V2.6-Flash (Xiaomi, $0.14 / $0.28).