Skip to main content
AI Choice EngineAI Choice Engine

Inception

Inception Mercury 2

Inception Mercury 2 is a Inception model with 128K context, $0.25 in / $0.75 out per million tokens, last verified September 2, 2026. Open weights: no. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.

Mercury 2 is Inception's diffusion-based LLM — radically fast (~1,000 tok/s) rather than autoregressive.

Catalog record checked September 2, 2026Individual provider fields may change

AI model specification details
SpecificationInception Mercury 2
ProviderInception
TierBalanced
Context window128K
Max outputNot verifiedUnverified
Input / 1M tokens$0.25
Output / 1M tokens$0.75
WeightsClosed
Parametersdiffusion LLM
Reasoning levelslow, medium, high
Modalitiestext
API model idmercury-2
ReleasedFebruary 24, 2026

Pricing tiers: $0.25/$0.75 per MTok (closed API). Diffusion-based 'dLLM' — ~1,000 tok/s.

Verified evidence

Published benchmark results

Each result keeps its source and measurement date visible. A missing benchmark is not treated as a zero.

Published benchmark results for Inception Mercury 2
BenchmarkScoreMeasuredSource
Artificial Analysis Intelligence Index382026-08-14Artificial Analysis · View source

Best for

  • Ultra-low-latency
  • High-throughput generation
  • Structured output

Watch out

Diffusion architecture; behaviour differs from autoregressive models.

Source receipts

Catalog figures for this model were checked against the following sources.

Common questions

Inception Mercury 2

Answered from the verified figures on this page rather than general guidance.

How much does Inception Mercury 2 cost per million tokens?

Inception Mercury 2 is listed at $0.25 per million input tokens and $0.75 per million output tokens at standard rates. Output tokens usually dominate real bills, so weigh the output rate more heavily than the input rate. $0.25/$0.75 per MTok (closed API). Diffusion-based 'dLLM' — ~1,000 tok/s.

What is Inception Mercury 2's context window?

Inception Mercury 2 accepts about 128K tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.

What is Inception Mercury 2 best for?

Inception Mercury 2 is a balanced tier from Inception. It suits ultra-low-latency, high-throughput generation, structured output. Diffusion architecture; behaviour differs from autoregressive models.