Skip to main content
TokenCost logoTokenCost

Mercury 2

Last verified September 7, 2026 · Inception Labs pricing ↗

$0.250/1M input · $0.750/1M output · 128K context · Inception Labs

Count Tokens

—
Tokens
—
Input Cost
—
Output Cost

Estimate Monthly Cost

Monthly Cost Estimator

Quick:
<$0.0001/mo

Pick a preset above or enter custom usage

Alternatives to Mercury 2

Pricing Details

Inception LabsFirst diffusion-based reasoning LLM. $0.25/$0.75 confirmed on three Inception surfaces that agree: the models page, the docs pricing table and its own /v1/models endpoint. Cached input $0.025/1M; cache writes are free ($0). Still sold, and as of September 7, 2026 it is the only model Inception's own API serves. Speed has three live figures and they are not interchangeable: 1,009 tok/s on NVIDIA Blackwell per Inception's launch post (vendor lab claim), 709.9 output tok/s measured by Artificial Analysis (2nd of 177), and 233 tok/s P50 across providers on OpenRouter's live traffic. We previously carried 788 tok/s attributed to Artificial Analysis; that page now reads 709.9 and we could not reconstruct the earlier capture, so this row is corrected to the current figure and dated. AA Intelligence Index 15. AIME 2025 91.1, GPQA Diamond 73.6, HumanEval 89.4 circulate widely but we could not find them on any Inception first-party page, so treat them as unverified. Verbosity caveat: produces ~2.6x output tokens of median model on AA suite. Sibling Mercury Edit 2 (FIM/NextEdit, 32K, 8,192 max output) shares the same rate card. Mercury Coder is absent from Inception's site, docs and API and from OpenRouter's catalogue. Estimated tokens (different tokenizer).
Input / 1M tokens
$0.25
Output / 1M tokens
$0.75
Context Window
128K
Max Output
50K

Price History

No price change recorded since 2026-08-11, when this site started tracking this model; it launched on 2026-05-12.

Frequently Asked Questions

Common questions about Mercury 2 pricing and usage

Read More About Mercury 2