Skip to main content
AI Choice Engine

Xiaomi

Xiaomi MiMo-V2.6-Flash

Xiaomi MiMo-V2.6-Flash is a Xiaomi model with 1.05M context, $0.14 in / $0.28 out per million tokens, last verified September 26, 2026. Open weights: yes. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.

MiMo-V2.6-Flash is Xiaomi's low-cost open-weight omni-modal MoE with a 1M context at $0.14/$0.28.

Catalog record checked September 26, 2026Individual provider fields may change

AI model specification details
SpecificationXiaomi MiMo-V2.6-Flash
ProviderXiaomi
TierBalanced
Context window1.05M
Max outputNot verifiedUnverified
Input / 1M tokens$0.14
Output / 1M tokens$0.28
WeightsOpen
Parameters~310B total / 15B active (MoE)
Reasoning levelsNot verifiedUnverified
Modalitiestext, image, video, audio
API model idmimo-v2.6-flash
ReleasedSeptember 21, 2026

Pricing tiers: $0.14/$0.28 per MTok (cache hit $0.0028) — unchanged from V2.5. Open weights; about 310B total / 15B active MoE.

Best for

  • Cheap omni-modal input
  • Open-weight deployments
  • High-volume long-context work

Watch out

No independent benchmark rows yet; licence and max output not re-verified for the Flash checkpoint.

Source receipts

Catalog figures for this model were checked against the following sources.

Common questions

Xiaomi MiMo-V2.6-Flash

Answered from the verified figures on this page rather than general guidance.

How much does Xiaomi MiMo-V2.6-Flash cost per million tokens?
Xiaomi MiMo-V2.6-Flash is listed at $0.14 per million input tokens and $0.28 per million output tokens at standard rates. Output tokens usually dominate real bills, so weigh the output rate more heavily than the input rate. $0.14/$0.28 per MTok (cache hit $0.0028) — unchanged from V2.5. Open weights; about 310B total / 15B active MoE.
What is Xiaomi MiMo-V2.6-Flash's context window?
Xiaomi MiMo-V2.6-Flash accepts about 1.05M tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.
What is Xiaomi MiMo-V2.6-Flash best for?
Xiaomi MiMo-V2.6-Flash is a balanced tier from Xiaomi. It suits cheap omni-modal input, open-weight deployments, high-volume long-context work. No independent benchmark rows yet; licence and max output not re-verified for the Flash checkpoint.
Can I self-host Xiaomi MiMo-V2.6-Flash?
Xiaomi MiMo-V2.6-Flash publishes open weights, but self-hosting depends on the licence, hardware footprint, quantisation quality, and serving stack. A hosted API is often cheaper until you have measured throughput and concurrency on your own hardware.