gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license.
Prices update automatically — checked daily against provider list prices.
Independent benchmark scores for gpt-oss-20b (batch), measured by Artificial Analysis. Higher is better. Measurement mode: High.
gpt-oss-20b (batch) is an open-weight AI model by OpenAI. You can download and self-host it for free; the prices below are hosted-API list prices, tracked daily, for when you prefer convenience over self-hosting.
gpt-oss-20b (batch) is an AI language model from OpenAI. It is open-weight: you can download it and run it on your own hardware, for free. It scores 9 on the Artificial Analysis intelligence index.
The weights are free and open — you can self-host gpt-oss-20b (batch) and pay nothing per token. If you prefer a hosted API, list prices are $0.02 per million input tokens and $0.11 per million output tokens.
Independent benchmarks from Artificial Analysis give it GPQA 68.8%, MMLU-Pro 74.8%, Humanity's Last Exam 11%, Long Context Reasoning 34.7%, LiveCodeBench 77.7%, SciCode 38.9%, AIME 2025 89.3%, IFBench 65.1%, τ²-Bench 60.2%, τ-Bench Banking 7%, Terminal-Bench 13.9%, Terminal-Bench Hard 10.6%. It is particularly used for code generation.
It generates about 202.6 tokens per second, with a median 0.44s delay before the first token. Measured independently by Artificial Analysis.
Yes. gpt-oss-20b (batch) has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.