Mercury 2
Last verified September 7, 2026 · Inception Labs pricing ↗$0.250/1M input · $0.750/1M output · 128K context · Inception Labs
Count Tokens
—
Tokens
—
Input Cost
—
Output Cost
Estimate Monthly Cost
Monthly Cost Estimator
Quick:
<$0.0001/mo
Pick a preset above or enter custom usage
Alternatives to Mercury 2
Cheaper Alternative
Qwen3 Coder Next
$0.120 in · $0.800 out · Alibaba
8% cheaper with same or larger context
Premium UpgradeNemotron 3 Super 120B
$0.300 in · $0.800 out · NVIDIA
Upgrade with 1000K context
Same ProviderMercury 2.5 (Preview)
$0.200 in · $0.750 out · Inception Labs
5% cheaper in Inception Labs ecosystem
Pricing Details
Inception LabsFirst diffusion-based reasoning LLM. $0.25/$0.75 confirmed on three Inception surfaces that agree: the models page, the docs pricing table and its own /v1/models endpoint. Cached input $0.025/1M; cache writes are free ($0). Still sold, and as of September 7, 2026 it is the only model Inception's own API serves. Speed has three live figures and they are not interchangeable: 1,009 tok/s on NVIDIA Blackwell per Inception's launch post (vendor lab claim), 709.9 output tok/s measured by Artificial Analysis (2nd of 177), and 233 tok/s P50 across providers on OpenRouter's live traffic. We previously carried 788 tok/s attributed to Artificial Analysis; that page now reads 709.9 and we could not reconstruct the earlier capture, so this row is corrected to the current figure and dated. AA Intelligence Index 15. AIME 2025 91.1, GPQA Diamond 73.6, HumanEval 89.4 circulate widely but we could not find them on any Inception first-party page, so treat them as unverified. Verbosity caveat: produces ~2.6x output tokens of median model on AA suite. Sibling Mercury Edit 2 (FIM/NextEdit, 32K, 8,192 max output) shares the same rate card. Mercury Coder is absent from Inception's site, docs and API and from OpenRouter's catalogue. Estimated tokens (different tokenizer).
Input / 1M tokens
$0.25
Output / 1M tokens
$0.75
Context Window
128K
Max Output
50K
Price History
No price change recorded since 2026-08-11, when this site started tracking this model; it launched on 2026-05-12.
Frequently Asked Questions
Common questions about Mercury 2 pricing and usage
Read More About Mercury 2
At 07:00 UTC tomorrow the cheapest model we can price goes up 5.00x, and the number that replaces it is not new. It is the card Inception has charged for Mercury 2 since March, with a fifth off the input line and nothing off the output line.
September 7, 2026 · 12 min read
Xiaomi MiMo UltraSpeed charges 3x for 10x the speed. The catch is you can only rent it for two weeks.
June 9, 2026 · 8 min read
Mercury 2 outputs at 788 tokens per second for $0.75 per million. The diffusion math turns frontier reasoning pricing into a rounding error.
May 13, 2026 · 11 min read