OpenAI · all models ›

OGPT-5.1API

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5.

400KContext window · tokens
$1.25Input price · per M tokens
$10Output price · per M tokens
OpenAIProvider

Prices update automatically — checked daily against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for GPT-5.1, measured by Artificial Analysis. Higher is better. Measurement mode: high.

Intelligence index37.5
Coding index49.4
Math index94
GPQA87.3%
MMLU-Pro87%
Humanity's Last Exam28.5%
Long Context Reasoning76.7%
LiveCodeBench86.8%
SciCode43.3%
AIME 202594%
IFBench72.9%
τ²-Bench81.9%
τ-Bench Banking15.9%
Terminal-Bench52.4%
Terminal-Bench Hard45.5%
💰 Blended price$3.438 / 1M tokens
📈 Value10.9 intelligence points per $
vs. models measured heretop 33%
Scores higher than 67% of the 273 models measured by Artificial Analysis and tracked here.
Models at this level cost $1.56 per 1M tokens (median of 25) — this one costs $3.44.
Cheaper and better on this index: Grok 4.6 · GPT-5.6 Sol (batch) · GLM 5.3 and 56 more
Benchmark data by Artificial Analysis

About this model

GPT-5.1 is a commercial AI model by OpenAI. The specifications below are tracked automatically: pricing is refreshed daily from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is GPT-5.1?

GPT-5.1 is an AI language model from OpenAI. It is a proprietary model, available through an API. It scores 37.5 on the Artificial Analysis intelligence index.

Is GPT-5.1 free?

GPT-5.1 is not free: it costs $1.25 per million input tokens and $10 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is GPT-5.1 good at?

Independent benchmarks from Artificial Analysis give it GPQA 87.3%, MMLU-Pro 87%, Humanity's Last Exam 28.5%, Long Context Reasoning 76.7%, LiveCodeBench 86.8%, SciCode 43.3%, AIME 2025 94%, IFBench 72.9%, τ²-Bench 81.9%, τ-Bench Banking 15.9%, Terminal-Bench 52.4%, Terminal-Bench Hard 45.5%. It is particularly used for code generation.

Related models

o1-proOpenAIo1-pro (batch)OpenAIGPT-5.5 ProOpenAIGPT-5.4 ProOpenAIGPT-5.2 ProOpenAIGPT-5 ProOpenAI