OpenAI · all models ›

OGPT-4.1API

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning.

1MContext window · tokens
$2Input price · per M tokens
$8Output price · per M tokens
OpenAIProvider

Prices update automatically — checked daily against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for GPT-4.1, measured by Artificial Analysis. Higher is better.

Intelligence index19.6
Math index34.7
GPQA66.6%
MMLU-Pro80.6%
Humanity's Last Exam4.2%
Long Context Reasoning64.3%
LiveCodeBench45.7%
SciCode38.1%
MATH-50091.3%
AIME43.7%
AIME 202534.7%
IFBench43%
τ²-Bench47.1%
Terminal-Bench Hard13.6%
💰 Blended price$3.5 / 1M tokens
📈 Value5.6 intelligence points per $
vs. models measured heretop 66%
Scores higher than 34% of the 273 models measured by Artificial Analysis and tracked here.
Models at this level cost $0.88 per 1M tokens (median of 23) — this one costs $3.5.
Cheaper and better on this index: Grok 4.6 · GPT-5.6 Sol (batch) · GLM 5.3 and 127 more
Benchmark data by Artificial Analysis

About this model

GPT-4.1 is a commercial AI model by OpenAI. The specifications below are tracked automatically: pricing is refreshed daily from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is GPT-4.1?

GPT-4.1 is an AI language model from OpenAI. It is a proprietary model, available through an API. It scores 19.6 on the Artificial Analysis intelligence index.

Is GPT-4.1 free?

GPT-4.1 is not free: it costs $2 per million input tokens and $8 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is GPT-4.1 good at?

Independent benchmarks from Artificial Analysis give it GPQA 66.6%, MMLU-Pro 80.6%, Humanity's Last Exam 4.2%, Long Context Reasoning 64.3%, LiveCodeBench 45.7%, SciCode 38.1%, MATH-500 91.3%, AIME 43.7%, AIME 2025 34.7%, IFBench 43%, τ²-Bench 47.1%, Terminal-Bench Hard 13.6%. It is particularly used for mathematical reasoning.

Related models

o1-proOpenAIo1-pro (batch)OpenAIGPT-5.5 ProOpenAIGPT-5.4 ProOpenAIGPT-5.2 ProOpenAIGPT-5 ProOpenAI