Xiaomi MiMo-V2.6-Flash is a Xiaomi model with 1.05M context, $0.14 in / $0.28 out per million tokens, last verified September 26, 2026. Open weights: yes. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.
MiMo-V2.6-Flash is Xiaomi's low-cost open-weight omni-modal MoE with a 1M context at $0.14/$0.28.
Catalog record checked September 26, 2026·Individual provider fields may change
AI model specification details
Specification
Xiaomi MiMo-V2.6-Flash
Provider
Xiaomi
Tier
Balanced
Context window
1.05M
Max output
Not verifiedUnverified
Input / 1M tokens
$0.14
Output / 1M tokens
$0.28
Weights
Open
Parameters
~310B total / 15B active (MoE)
Reasoning levels
Not verifiedUnverified
Modalities
text, image, video, audio
API model id
mimo-v2.6-flash
Released
September 21, 2026
Pricing tiers: $0.14/$0.28 per MTok (cache hit $0.0028) — unchanged from V2.5. Open weights; about 310B total / 15B active MoE.
Best for
Cheap omni-modal input
Open-weight deployments
High-volume long-context work
Watch out
No independent benchmark rows yet; licence and max output not re-verified for the Flash checkpoint.
Source receipts
Catalog figures for this model were checked against the following sources.
Answered from the verified figures on this page rather than general guidance.
How much does Xiaomi MiMo-V2.6-Flash cost per million tokens?
Xiaomi MiMo-V2.6-Flash is listed at $0.14 per million input tokens and $0.28 per million output tokens at standard rates. Output tokens usually dominate real bills, so weigh the output rate more heavily than the input rate. $0.14/$0.28 per MTok (cache hit $0.0028) — unchanged from V2.5. Open weights; about 310B total / 15B active MoE.
What is Xiaomi MiMo-V2.6-Flash's context window?
Xiaomi MiMo-V2.6-Flash accepts about 1.05M tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.
What is Xiaomi MiMo-V2.6-Flash best for?
Xiaomi MiMo-V2.6-Flash is a balanced tier from Xiaomi. It suits cheap omni-modal input, open-weight deployments, high-volume long-context work. No independent benchmark rows yet; licence and max output not re-verified for the Flash checkpoint.
Can I self-host Xiaomi MiMo-V2.6-Flash?
Xiaomi MiMo-V2.6-Flash publishes open weights, but self-hosting depends on the licence, hardware footprint, quantisation quality, and serving stack. A hosted API is often cheaper until you have measured throughput and concurrency on your own hardware.