Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows.
Prices update automatically — checked daily against provider list prices.
Independent benchmark scores for Claude Opus 4, measured by Artificial Analysis. Higher is better. Measurement mode: Reasoning.
Claude Opus 4 is a commercial AI model by Anthropic. The specifications below are tracked automatically: pricing is refreshed daily from public list prices, so the numbers on this page reflect the current cost of using the model through its API.
Claude Opus 4 is an AI language model from Anthropic. It is a proprietary model, available through an API. It scores 31.7 on the Artificial Analysis intelligence index.
Claude Opus 4 is not free: it costs $15 per million input tokens and $75 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.
Independent benchmarks from Artificial Analysis give it GPQA 79.6%, MMLU-Pro 87.3%, Humanity's Last Exam 12.3%, Long Context Reasoning 36.3%, LiveCodeBench 63.6%, SciCode 39.8%, MATH-500 98.2%, AIME 75.7%, AIME 2025 73.3%, IFBench 53.7%, τ²-Bench 73.4%, Terminal-Bench Hard 31.1%. It is particularly used for mathematical reasoning.