Inception
Inception Mercury 2
Inception Mercury 2 is a Inception model with 128K context, $0.25 in / $0.75 out per million tokens, last verified September 2, 2026. Open weights: no. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.
Mercury 2 is Inception's diffusion-based LLM — radically fast (~1,000 tok/s) rather than autoregressive.
Catalog record checked September 2, 2026Individual provider fields may change
| Specification | Inception Mercury 2 |
|---|---|
| Provider | Inception |
| Tier | Balanced |
| Context window | 128K |
| Max output | Not verifiedUnverified |
| Input / 1M tokens | $0.25 |
| Output / 1M tokens | $0.75 |
| Weights | Closed |
| Parameters | diffusion LLM |
| Reasoning levels | low, medium, high |
| Modalities | text |
| API model id | mercury-2 |
| Released | February 24, 2026 |
Pricing tiers: $0.25/$0.75 per MTok (closed API). Diffusion-based 'dLLM' — ~1,000 tok/s.
Verified evidence
Published benchmark results
Each result keeps its source and measurement date visible. A missing benchmark is not treated as a zero.
| Benchmark | Score | Measured | Source |
|---|---|---|---|
| Artificial Analysis Intelligence Index | 38 | 2026-08-14 | Artificial Analysis · View source |
Best for
- Ultra-low-latency
- High-throughput generation
- Structured output
Watch out
Source receipts
Catalog figures for this model were checked against the following sources.
- Inception Labs — Mercury 2 (accessed 2026-08-29)
- BusinessWire — Mercury 2 (accessed 2026-08-29)
Common questions
Inception Mercury 2
Answered from the verified figures on this page rather than general guidance.
How much does Inception Mercury 2 cost per million tokens?
Inception Mercury 2 is listed at $0.25 per million input tokens and $0.75 per million output tokens at standard rates. Output tokens usually dominate real bills, so weigh the output rate more heavily than the input rate. $0.25/$0.75 per MTok (closed API). Diffusion-based 'dLLM' — ~1,000 tok/s.
What is Inception Mercury 2's context window?
Inception Mercury 2 accepts about 128K tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.
What is Inception Mercury 2 best for?
Inception Mercury 2 is a balanced tier from Inception. It suits ultra-low-latency, high-throughput generation, structured output. Diffusion architecture; behaviour differs from autoregressive models.