Anthropic · all models ›

AClaude Sonnet 4API

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability.

1MContext window · tokens
$3Input price · per M tokens
$15Output price · per M tokens
AnthropicProvider

Prices update automatically — checked daily against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for Claude Sonnet 4, measured by Artificial Analysis. Higher is better. Measurement mode: Reasoning.

Intelligence index29.8
Coding index37.6
Math index74.3
GPQA77.7%
MMLU-Pro84.2%
Humanity's Last Exam10.7%
Long Context Reasoning69.7%
LiveCodeBench65.5%
SciCode40%
MATH-50099.1%
AIME77.3%
AIME 202574.3%
IFBench54.7%
τ²-Bench64.6%
τ-Bench Banking16.7%
Terminal-Bench36.3%
Terminal-Bench Hard31.1%
💰 Blended price$6 / 1M tokens
📈 Value5 intelligence points per $
vs. models measured heretop 49%
Scores higher than 51% of the 273 models measured by Artificial Analysis and tracked here.
Models at this level cost $0.69 per 1M tokens (median of 14) — this one costs $6.
Cheaper and better on this index: Claude Opus 5 (batch) · GPT-5.6 Sol · Grok 4.6 and 106 more
Benchmark data by Artificial Analysis

About this model

Claude Sonnet 4 is a commercial AI model by Anthropic. The specifications below are tracked automatically: pricing is refreshed daily from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is Claude Sonnet 4?

Claude Sonnet 4 is an AI language model from Anthropic. It is a proprietary model, available through an API. It scores 29.8 on the Artificial Analysis intelligence index.

Is Claude Sonnet 4 free?

Claude Sonnet 4 is not free: it costs $3 per million input tokens and $15 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is Claude Sonnet 4 good at?

Independent benchmarks from Artificial Analysis give it GPQA 77.7%, MMLU-Pro 84.2%, Humanity's Last Exam 10.7%, Long Context Reasoning 69.7%, LiveCodeBench 65.5%, SciCode 40%, MATH-500 99.1%, AIME 77.3%, AIME 2025 74.3%, IFBench 54.7%, τ²-Bench 64.6%, τ-Bench Banking 16.7%, Terminal-Bench 36.3%, Terminal-Bench Hard 31.1%. It is particularly used for code generation.

Related models

Claude Opus 4.7 (Fast)AnthropicClaude Opus 4.1AnthropicClaude Opus 4AnthropicClaude Fable 5AnthropicClaude Opus 5 (Fast)AnthropicClaude Opus 4.8 (Fast)Anthropic