Skip to main content
AI Choice EngineAI Choice Engine

Qwen

Qwen 3.8 Flash Next

Qwen 3.8 Flash Next is a Qwen model with 262K context, Not verified in / Not verified out per million tokens, last verified September 3, 2026. Open weights: yes. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.

Qwen 3.8 Flash Next is Alibaba's experimental preview of the Qwen 4 architecture — very low active parameters plus an unusual n-gram embedding component for cheap long-context.

Catalog record checked September 3, 2026Individual provider fields may change

AI model specification details
SpecificationQwen 3.8 Flash Next
ProviderQwen
TierBudget
Context window262K
Max outputNot verifiedUnverified
Input / 1M tokensNot verifiedUnverified
Output / 1M tokensNot verifiedUnverified
WeightsOpen
Parameters125B total / 6B active (MoE) + 51B n-gram embedding table
Reasoning levelsNot verifiedUnverified
Modalitiestext, image, video
Licenseqwen-community-1.0
API model idqwen3.8-flash
ReleasedAugust 26, 2026

Pricing tiers: QwenCloud reportedly prices it at $0.16/$0.47 per MTok (single source as of 2026-08-29 — left unverified). Native 262K context, extensible to 1M.

Verified evidence

Published benchmark results

Each result keeps its source and measurement date visible. A missing benchmark is not treated as a zero.

Published benchmark results for Qwen 3.8 Flash Next
BenchmarkScoreMeasuredSource
Artificial Analysis Intelligence Index562026-08-28Artificial Analysis · View source

Best for

  • Early testing of Qwen 4 architecture
  • Cheap high-throughput work
  • Multimodal input

Watch out

Experimental preview; qwen-community-1.0 license is more restrictive than MIT/Apache; benchmark claims are vendor-reported.

Source receipts

Catalog figures for this model were checked against the following sources.

Common questions

Qwen 3.8 Flash Next

Answered from the verified figures on this page rather than general guidance.

What is Qwen 3.8 Flash Next's context window?

Qwen 3.8 Flash Next accepts about 262K tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.

What is Qwen 3.8 Flash Next best for?

Qwen 3.8 Flash Next is a budget tier from Qwen. It suits early testing of qwen 4 architecture, cheap high-throughput work, multimodal input. Experimental preview; qwen-community-1.0 license is more restrictive than MIT/Apache; benchmark claims are vendor-reported.

Can I self-host Qwen 3.8 Flash Next?

Qwen 3.8 Flash Next publishes open weights, but self-hosting depends on the licence, hardware footprint, quantisation quality, and serving stack. A hosted API is often cheaper until you have measured throughput and concurrency on your own hardware.