Anthropic · all models ›

AClaude Opus 4.1API

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks.

200KContext window · tokens
$15Input price · per M tokens
$75Output price · per M tokens
AnthropicProvider

Prices update automatically — checked daily against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for Claude Opus 4.1, measured by Artificial Analysis. Higher is better. Measurement mode: Reasoning.

Intelligence index34.5
Math index80.3
GPQA80.9%
MMLU-Pro88%
Humanity's Last Exam12.5%
Long Context Reasoning73.3%
LiveCodeBench65.4%
SciCode40.9%
AIME 202580.3%
IFBench55.4%
τ²-Bench71.4%
Terminal-Bench Hard34.3%
💰 Blended price$30 / 1M tokens
📈 Value1.2 intelligence points per $
vs. models measured heretop 40%
Scores higher than 60% of the 273 models measured by Artificial Analysis and tracked here.
Models at this level cost $1.2 per 1M tokens (median of 21) — this one costs $30.
Cheaper and better on this index: Claude Opus 5 (Fast) · Claude Opus 5 · Claude Opus 5 (batch) and 97 more
Benchmark data by Artificial Analysis

About this model

Claude Opus 4.1 is a commercial AI model by Anthropic. The specifications below are tracked automatically: pricing is refreshed daily from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is Claude Opus 4.1?

Claude Opus 4.1 is an AI language model from Anthropic. It is a proprietary model, available through an API. It scores 34.5 on the Artificial Analysis intelligence index.

Is Claude Opus 4.1 free?

Claude Opus 4.1 is not free: it costs $15 per million input tokens and $75 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is Claude Opus 4.1 good at?

Independent benchmarks from Artificial Analysis give it GPQA 80.9%, MMLU-Pro 88%, Humanity's Last Exam 12.5%, Long Context Reasoning 73.3%, LiveCodeBench 65.4%, SciCode 40.9%, AIME 2025 80.3%, IFBench 55.4%, τ²-Bench 71.4%, Terminal-Bench Hard 34.3%. It is particularly used for mathematical reasoning.

Related models

Claude Opus 4.7 (Fast)AnthropicClaude Opus 4AnthropicClaude Fable 5AnthropicClaude Opus 5 (Fast)AnthropicClaude Opus 4.8 (Fast)AnthropicClaude Opus 4.1 (batch)Anthropic