Inception · all models ›

IMercury 2API

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM).

128KContext window · tokens
$0.25Input price · per M tokens
$0.75Output price · per M tokens
InceptionProvider

Prices update automatically — checked daily against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for Mercury 2, measured by Artificial Analysis. Higher is better.

Intelligence index21.9
Coding index31.1
GPQA77%
Humanity's Last Exam17.1%
Long Context Reasoning40.7%
SciCode38.7%
IFBench69.8%
τ²-Bench70.8%
τ-Bench Banking9.5%
Terminal-Bench27.3%
Terminal-Bench Hard26.5%
⚡ Speed1204.9 tokens/sec
⏱ Latency3.71s to first token
💰 Blended price$0.375 / 1M tokens
📈 Value58.4 intelligence points per $
vs. models measured heretop 60%
Scores higher than 40% of the 273 models measured by Artificial Analysis and tracked here.
Models at this level cost $0.76 per 1M tokens (median of 16) — this one costs $0.38.
Cheaper and better on this index: GLM 5.3 Flash · GPT-5.6 Luna (batch) · DeepSeek V4 Flash 0731 and 19 more
Benchmark data by Artificial Analysis

About this model

Mercury 2 is a commercial AI model by Inception. The specifications below are tracked automatically: pricing is refreshed daily from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is Mercury 2?

Mercury 2 is an AI language model from Inception. It is a proprietary model, available through an API. It scores 21.9 on the Artificial Analysis intelligence index.

Is Mercury 2 free?

Mercury 2 is not free: it costs $0.25 per million input tokens and $0.75 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is Mercury 2 good at?

Independent benchmarks from Artificial Analysis give it GPQA 77%, Humanity's Last Exam 17.1%, Long Context Reasoning 40.7%, SciCode 38.7%, IFBench 69.8%, τ²-Bench 70.8%, τ-Bench Banking 9.5%, Terminal-Bench 27.3%, Terminal-Bench Hard 26.5%. It is particularly used for code generation.

How fast is Mercury 2?

It generates about 1204.9 tokens per second, with a median 3.71s delay before the first token. Measured independently by Artificial Analysis.

Related models