Skip to main content
AI Choice Engine

Qwen

Qwen 3.8 Omni Flash

Qwen 3.8 Omni Flash is a Qwen model with 1M context, $0.15 in / $0.47 out per million tokens, last verified September 26, 2026. Open weights: no. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.

Qwen 3.8 Omni Flash is Alibaba's low-cost omni-modal API model — text, image, audio and video input with a 1M-token context.

Catalog record checked September 26, 2026Individual provider fields may change

AI model specification details
SpecificationQwen 3.8 Omni Flash
ProviderQwen
TierBudget
Context window1M
Max outputNot verifiedUnverified
Input / 1M tokens$0.15
Output / 1M tokens$0.47
WeightsClosed
ParametersNot disclosed
Reasoning levelsNot verifiedUnverified
Modalitiestext, image, audio, video
API model idqwen3.8-omni-flash
ReleasedSeptember 18, 2026

Pricing tiers: $0.15/$0.47 per MTok for text, image and video input on Alibaba Cloud Model Studio (cache hit $0.016); audio input is priced separately. A -realtime variant launched 2026-09-21.

Best for

  • Cheap multimodal input
  • Audio and video understanding
  • Voice agents (realtime variant)

Watch out

API-only as far as verified (no open-weight release confirmed); max output is not verified. No independent benchmark rows yet.

History

Qwen 3.8 Omni Flash timeline

Release and status events for this model that passed two-source verification — dated, sourced, and linked.

  • Qwen3.8-Omni-FlashReleaseSeptember 18, 2026

    Alibaba adds Qwen3.8-Omni-Flash to Model Studio: text, image, audio and video input with a 1M-token context at $0.15/$0.47 per MTok. A -realtime variant followed on 2026-09-21.

    QwenCloud model changelog

Source receipts

Catalog figures for this model were checked against the following sources.

Common questions

Qwen 3.8 Omni Flash

Answered from the verified figures on this page rather than general guidance.

How much does Qwen 3.8 Omni Flash cost per million tokens?
Qwen 3.8 Omni Flash is listed at $0.15 per million input tokens and $0.47 per million output tokens at standard rates. Output tokens usually dominate real bills, so weigh the output rate more heavily than the input rate. $0.15/$0.47 per MTok for text, image and video input on Alibaba Cloud Model Studio (cache hit $0.016); audio input is priced separately. A -realtime variant launched 2026-09-21.
What is Qwen 3.8 Omni Flash's context window?
Qwen 3.8 Omni Flash accepts about 1M tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.
What is Qwen 3.8 Omni Flash best for?
Qwen 3.8 Omni Flash is a budget tier from Qwen. It suits cheap multimodal input, audio and video understanding, voice agents (realtime variant). API-only as far as verified (no open-weight release confirmed); max output is not verified. No independent benchmark rows yet.