OpenAI · all models ›

OGPT-5.1-CodexAPI

GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows.

400KContext window · tokens
$1.25Input price · per M tokens
$10Output price · per M tokens
OpenAIProvider

Prices update automatically — checked daily against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for GPT-5.1-Codex, measured by Artificial Analysis. Higher is better. Measurement mode: high.

Intelligence index35.6
Math index95.7
GPQA86%
MMLU-Pro86%
Humanity's Last Exam25.7%
Long Context Reasoning69%
LiveCodeBench84.9%
SciCode40.2%
AIME 202595.7%
IFBench70%
τ²-Bench83%
Terminal-Bench Hard34.8%
💰 Blended price$3.438 / 1M tokens
📈 Value10.4 intelligence points per $
vs. models measured heretop 37%
Scores higher than 63% of the 273 models measured by Artificial Analysis and tracked here.
Models at this level cost $1.72 per 1M tokens (median of 25) — this one costs $3.44.
Cheaper and better on this index: Grok 4.6 · GPT-5.6 Sol (batch) · GLM 5.3 and 63 more
Benchmark data by Artificial Analysis

About this model

GPT-5.1-Codex is a commercial AI model by OpenAI. The specifications below are tracked automatically: pricing is refreshed daily from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is GPT-5.1-Codex?

GPT-5.1-Codex is an AI language model from OpenAI. It is a proprietary model, available through an API. It scores 35.6 on the Artificial Analysis intelligence index.

Is GPT-5.1-Codex free?

GPT-5.1-Codex is not free: it costs $1.25 per million input tokens and $10 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is GPT-5.1-Codex good at?

Independent benchmarks from Artificial Analysis give it GPQA 86%, MMLU-Pro 86%, Humanity's Last Exam 25.7%, Long Context Reasoning 69%, LiveCodeBench 84.9%, SciCode 40.2%, AIME 2025 95.7%, IFBench 70%, τ²-Bench 83%, Terminal-Bench Hard 34.8%. It is particularly used for mathematical reasoning.

Related models

o1-proOpenAIo1-pro (batch)OpenAIGPT-5.5 ProOpenAIGPT-5.4 ProOpenAIGPT-5.2 ProOpenAIGPT-5 ProOpenAI