Back to models
Alibaba · Qwen

Qwen3.8 Omni Flash

qwen3.8-omni-flash

Alibaba's flagship Qwen reasoning model with a 1M-token context and text, image and video input.

Compare models
Input and output modalities
Input/output price
$0.12 / $0.4 / 1M token
Context window
1M
Providers
1
Release date
Sep 21, 2026

Providers

Effective prices, cache hit rate, and live availability for every provider.

Alibaba Cloud
$0.12 / $0.4
-
-

Effective prices include each mapping's discount. Billing uses the provider mapping actually selected for the request.

Performance

DDSketch latency percentiles by minute or hour, defaulting to P95.

Latency

Time to first token

Availability

Cache hit rate

Applications

Client-reported applications ranked by token usage over the last 24 hours.

Last 24 hours
No application source data