Claude Fable 5.1 (batch) vs Mercury 2 compared — price per token, context window, multimodality, openness and which to choose. Full 2026 breakdown.
Claude Fable 5.1 (batch) — full profile › · Mercury 2 — full profile ›
Prices & specs refreshed from live data · olud.ai
Open-model prices = cheapest provider via OpenRouter; official maker rates may be higher.
| Spec | Claude Fable 5.1 (batch) | Mercury 2 | Winner |
|---|---|---|---|
| Maker | Anthropic | Inception | – |
| Type | Proprietary | Proprietary | – |
| Context window | 1M tokens | 128K tokens | |
| Input price | $5/M | $0.25/M | INMercury 2 |
| Output price | $25/M | $0.75/M | INMercury 2 |
| Vision / multimodal | Yes | No | |
| Tool / function calling | Yes | Yes | = Tie |
| Self-hostable | No (API only) | No (API only) | – |
| License | Proprietary | Proprietary | – |
Mercury 2 is ~33× cheaper than Claude Fable 5.1 (batch) on output tokens ($0.75 vs $25 per M tokens).
| Capability | Claude Fable 5.1 (batch) | Mercury 2 |
|---|---|---|
| Open weights (downloadable) | ✗ | ✗ |
| Self-hostable | ✗ | ✗ |
| Runs fully offline | ✗ | ✗ |
| Vision / multimodal | ✓ | ✗ |
| Tool / function calling | ✓ | ✓ |
| 1M+ context window | ✓ | ✗ |
Independent benchmark scores measured by Artificial Analysis. Higher is better (except latency).
Benchmark data by Artificial Analysis.
| Criterion | Claude Fable 5.1 (batch) | Mercury 2 |
|---|---|---|
| Cost-efficiency | 3.0 | 5.0 |
| Context window | 5.0 | 3.5 |
| Openness | 1.5 | 1.5 |
| Self-hosting | 1.0 | 1.0 |
| Multimodality | 5.0 | 3.5 |
Scores come from live data — output price (cost), context length, open vs closed weights (openness & self-hosting) and vision/tool support (multimodality). Raw task quality isn't scored here; it depends on your benchmark — see the verdict.
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
Mercury 2 is cheaper on output ($0.75/M vs $25/M).
Claude Fable 5.1 (batch) offers the larger context window (1M tokens).
Choose Mercury 2 for the lower output price ($0.75/M vs $25/M). Choose Claude Fable 5.1 (batch) if you need the larger 1M context window.
Browse the full open-source model leaderboard, thousands of tools and live benchmarks — all in one place.
Open the leaderboard →