Qwen
Qwen 3.8 Flash Next
Qwen 3.8 Flash Next is a Qwen model with 262K context, Not verified in / Not verified out per million tokens, last verified September 3, 2026. Open weights: yes. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.
Qwen 3.8 Flash Next is Alibaba's experimental preview of the Qwen 4 architecture — very low active parameters plus an unusual n-gram embedding component for cheap long-context.
Catalog record checked September 3, 2026Individual provider fields may change
| Specification | Qwen 3.8 Flash Next |
|---|---|
| Provider | Qwen |
| Tier | Budget |
| Context window | 262K |
| Max output | Not verifiedUnverified |
| Input / 1M tokens | Not verifiedUnverified |
| Output / 1M tokens | Not verifiedUnverified |
| Weights | Open |
| Parameters | 125B total / 6B active (MoE) + 51B n-gram embedding table |
| Reasoning levels | Not verifiedUnverified |
| Modalities | text, image, video |
| License | qwen-community-1.0 |
| API model id | qwen3.8-flash |
| Released | August 26, 2026 |
Pricing tiers: QwenCloud reportedly prices it at $0.16/$0.47 per MTok (single source as of 2026-08-29 — left unverified). Native 262K context, extensible to 1M.
Verified evidence
Published benchmark results
Each result keeps its source and measurement date visible. A missing benchmark is not treated as a zero.
| Benchmark | Score | Measured | Source |
|---|---|---|---|
| Artificial Analysis Intelligence Index | 56 | 2026-08-28 | Artificial Analysis · View source |
Best for
- Early testing of Qwen 4 architecture
- Cheap high-throughput work
- Multimodal input
Watch out
Source receipts
Catalog figures for this model were checked against the following sources.
- The New Stack — Qwen3.8-Flash previews Qwen4 (accessed 2026-08-29)
- Yotta Labs — Qwen 3.8-Flash-Next specs (accessed 2026-08-29)
Common questions
Qwen 3.8 Flash Next
Answered from the verified figures on this page rather than general guidance.
What is Qwen 3.8 Flash Next's context window?
Qwen 3.8 Flash Next accepts about 262K tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.
What is Qwen 3.8 Flash Next best for?
Qwen 3.8 Flash Next is a budget tier from Qwen. It suits early testing of qwen 4 architecture, cheap high-throughput work, multimodal input. Experimental preview; qwen-community-1.0 license is more restrictive than MIT/Apache; benchmark claims are vendor-reported.
Can I self-host Qwen 3.8 Flash Next?
Qwen 3.8 Flash Next publishes open weights, but self-hosting depends on the licence, hardware footprint, quantisation quality, and serving stack. A hosted API is often cheaper until you have measured throughput and concurrency on your own hardware.