AI Models · Open-Source vs Open-Source

DeepSeek V4 Pro Open vs Mistral Large 3 2512 Open

DeepSeek V4 Pro vs Mistral Large 3 2512 compared — price per token, context window, multimodality, openness and which to choose. Full 2026 breakdown.

Prices & specs refreshed from live data · olud.ai

Open-model prices = cheapest provider via OpenRouter; official maker rates may be higher.

Choose DeepSeek V4 Pro for the lower output price ($0.87/M vs $1.5/M). It also gives you the larger 1M context window.

DeepSeek V4 Pro vs Mistral Large 3 2512 specs

SpecDeepSeek V4 ProMistral Large 3 2512
MakerDeepSeekMistral AI
TypeOpen-weightOpen-weight
Context window1M tokens262K tokens
Input price$0.44/M · free self-host$0.5/M · free self-host
Output price$0.87/M · free self-host$1.5/M · free self-host
Vision / multimodalNoYes
Tool / function callingYesYes
Self-hostableYesYes
LicenseMITApache 2.0

Feature comparison

CapabilityDeepSeek V4 ProMistral Large 3 2512
Open weights (downloadable)
Self-hostable
Runs fully offline
Vision / multimodal
Tool / function calling
1M+ context window

Benchmarks: DeepSeek V4 Pro vs Mistral Large 3 2512

Independent benchmark scores measured by Artificial Analysis. Higher is better (except latency).

DeepSeek V4 Pro delivers 3.8× more intelligence per dollar.
DeepSeek V4 Pro
Mistral Large 3 2512
Intelligence index
44.3
15.9
Coding index
59.4
20.1
GPQA
88.8%
68%
Humanity's Last Exam
35.9%
4.1%
Long Context Reasoning
66.3%
34.7%
SciCode
50%
36.2%
IFBench
76.5%
36.2%
τ²-Bench
96.2%
24.6%
τ-Bench Banking
25.8%
5.8%
Terminal-Bench
64%
12%
Terminal-Bench Hard
46.2%
15.9%
Math index
38
MMLU-Pro
80.7%
LiveCodeBench
46.5%
AIME 2025
38%
Speed
70.8 tok/s
60.9 tok/s
Latency
0.91s
1.05s
Intelligence per $
81.4
21.2

Benchmark data by Artificial Analysis.

How DeepSeek V4 Pro and Mistral Large 3 2512 score

🤝 Neck and neck on these criteria (4.7 vs 4.7 / 5).
CriterionDeepSeek V4 ProMistral Large 3 2512
Cost-efficiency5.04.5
Context window5.04.0
Openness5.05.0
Self-hosting5.05.0
Multimodality3.55.0

Scores come from live data — output price (cost), context length, open vs closed weights (openness & self-hosting) and vision/tool support (multimodality). Raw task quality isn't scored here; it depends on your benchmark — see the verdict.

What each model is

DeepSeek V4 Pro Open

DeepSeek · Open-weight

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

Mistral Large 3 2512 Open

Mistral AI · Open-weight

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

Other models in these families

These variants are tracked but not compared here — one page per family keeps the comparison readable.

Other variants tracked
DeepSeek V4 FlashDeepSeek V3.1 TerminusDeepSeek V3.2 ExpDeepSeek V3.2DeepSeek V3.1R1 0528R1DeepSeek V3 0324DeepSeek V3
Other variants tracked
Mistral Medium 3.5Mistral Small 4Devstral 2 2512Mistral Medium 3.1Mistral Small 3.1 24BMistral Medium 3Ministral 3 14B 2512Mistral Small 3.2 24BMinistral 3 8B 2512Mistral LargeMistral Large 2407Mistral Small 3SabaMinistral 3 3B 2512Mixtral 8x22B InstructCodestral 2508UncensoredVoxtral Small 24B 2507Mistral Nemo

Frequently asked questions

DeepSeek V4 Pro vs Mistral Large 3 2512 — which is cheaper?

DeepSeek V4 Pro is cheaper on output ($0.87/M vs $1.5/M).

Which has the larger context window?

DeepSeek V4 Pro offers the larger context window (1M tokens).

DeepSeek V4 Pro vs Mistral Large 3 2512 — which should I pick in 2026?

Choose DeepSeek V4 Pro for the lower output price ($0.87/M vs $1.5/M). It also gives you the larger 1M context window.

People also compare

Explore more open-source AI

Browse the full open-source model leaderboard, thousands of tools and live benchmarks — all in one place.

Open the leaderboard →