The best open-weight, self-hostable alternatives to Gemini 3.1 Pro Preview in 2026 — compared on price, context window and capabilities. Run them locally and cut API costs.
Refreshed from live data · OpenSourceAI.tech
Gemini 3.1 Pro Preview is a proprietary, API-only model. These open-weight models can be self-hosted, run offline and used at a fraction of the cost — here's how the top ones stack up.
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
DeepSeek V4 Flash vs Gemini 3.1 Pro Preview →GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
GLM 5.2 vs Gemini 3.1 Pro Preview →MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
MiniMax M3 vs Gemini 3.1 Pro Preview →NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Nemotron 3 Ultra (free) vs Gemini 3.1 Pro Preview →Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Kimi K3 vs Gemini 3.1 Pro Preview →gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimize
gpt-oss-120b vs Gemini 3.1 Pro Preview →Live ranking of open-weight models with pricing, context windows and capabilities.
Open the leaderboard →