AI Models · Open-Source Alternatives

Open-Source Alternatives to GPT-4

The best open-weight, self-hostable alternatives to GPT-4 in 2026 — compared on price, context window and capabilities. Run them locally and cut API costs.

Refreshed from live data · olud.ai

The best open-weight alternatives to GPT-4

GPT-4 is a proprietary, API-only model. These open-weight models can be self-hosted, run offline and used at a fraction of the cost — here's how the top ones stack up.

GLM 5.3 Open

Z.AI · 1.3M ctx · $4.4/M out

38.2 pts ABOVE GPT-4 on the Artificial Analysis intelligence index · 13.6× cheaper per million output tokens

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

GLM 5.3 vs GPT-4 →

Kimi K3 (batch) Open

Moonshot AI · 1M ctx · $15/M out

37.1 pts ABOVE GPT-4 on the Artificial Analysis intelligence index · 4.0× cheaper per million output tokens

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

Kimi K3 (batch) vs GPT-4 →

Qwen3.8 Max (0902) Open

Alibaba · 1M ctx · $6/M out

33.6 pts ABOVE GPT-4 on the Artificial Analysis intelligence index · 10.0× cheaper per million output tokens

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

Qwen3.8 Max (0902) vs GPT-4 →

DeepSeek V4.1 Flash Open

DeepSeek · 1M ctx · $0.6/M out

32.8 pts ABOVE GPT-4 on the Artificial Analysis intelligence index · 100.0× cheaper per million output tokens

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

DeepSeek V4.1 Flash vs GPT-4 →

MiniMax M3 Open

MiniMax · 1M ctx · $1.2/M out

22.9 pts ABOVE GPT-4 on the Artificial Analysis intelligence index · 50.0× cheaper per million output tokens

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

MiniMax M3 vs GPT-4 →

XI MiMo-V2.5-Pro Open

Xiaomi · 1.1M ctx · $0.87/M out

19.7 pts ABOVE GPT-4 on the Artificial Analysis intelligence index · 69.0× cheaper per million output tokens

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

MiMo-V2.5-Pro vs GPT-4 →

See the full open-source leaderboard

Live ranking of open-weight models with pricing, context windows and capabilities.

Open the leaderboard →