AI Models · Paid vs Paid

Gemini 3 Flash Preview Paid vs Gemini 2.5 Flash Paid

Gemini 3 Flash Preview vs Gemini 2.5 Flash compared — price per token, context window, multimodality, openness and which to choose. Full 2026 breakdown.

Prices & specs refreshed from live data · olud.ai

Open-model prices = cheapest provider via OpenRouter; official maker rates may be higher.

Choose Gemini 2.5 Flash for the lower output price ($2.5/M vs $3/M). Choose Gemini 3 Flash Preview if you need the larger 1M context window.

Gemini 3 Flash Preview vs Gemini 2.5 Flash specs

SpecGemini 3 Flash PreviewGemini 2.5 Flash
MakerGoogleGoogle
TypeProprietaryProprietary
Context window1M tokens1M tokens
Input price$0.5/M$0.3/M
Output price$3/M$2.5/M
Vision / multimodalYesYes
Tool / function callingYesYes
Self-hostableNo (API only)No (API only)
LicenseProprietaryProprietary

Feature comparison

CapabilityGemini 3 Flash PreviewGemini 2.5 Flash
Open weights (downloadable)
Self-hostable
Runs fully offline
Vision / multimodal
Tool / function calling
1M+ context window

Benchmarks: Gemini 3 Flash Preview vs Gemini 2.5 Flash

Independent benchmark scores measured by Artificial Analysis. Higher is better (except latency).

Gemini 3 Flash Preview delivers more intelligence per dollar.
Gemini 3 Flash Preview
Gemini 2.5 Flash
Intelligence index
37.8
14.1
Math index
97
60.3
GPQA
89.8%
68.3%
MMLU-Pro
89%
80.9%
Humanity's Last Exam
34.7%
5.1%
Long Context Reasoning
66.3%
45.9%
LiveCodeBench
90.8%
49.5%
SciCode
50.6%
29.1%
AIME 2025
97%
60.3%
IFBench
78%
39%
τ²-Bench
80.4%
14.9%
τ-Bench Banking
17.5%
Terminal-Bench Hard
38.6%
12.1%
MATH-500
93.2%
AIME
50%
Speed
0 tok/s
0 tok/s
Latency
0s
0s
Intelligence per $
33.6
16.6

Benchmark data by Artificial Analysis.

How Gemini 3 Flash Preview and Gemini 2.5 Flash score

🤝 Neck and neck on these criteria (3.3 vs 3.4 / 5).
CriterionGemini 3 Flash PreviewGemini 2.5 Flash
Cost-efficiency4.04.5
Context window5.05.0
Openness1.51.5
Self-hosting1.01.0
Multimodality5.05.0

Scores come from live data — output price (cost), context length, open vs closed weights (openness & self-hosting) and vision/tool support (multimodality). Raw task quality isn't scored here; it depends on your benchmark — see the verdict.

What each model is

Gemini 3 Flash Preview Paid

Google · Proprietary

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

Gemini 2.5 Flash Paid

Google · Proprietary

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with grea

Other models in these families

These variants are tracked but not compared here — one page per family keeps the comparison readable.

Other variants tracked
Gemini 3.5 FlashGemini 3.6 FlashGemini 2.5 Flash
Other variants tracked
Gemini 3.5 FlashGemini 3.6 FlashGemini 3 Flash Preview

Frequently asked questions

Gemini 3 Flash Preview vs Gemini 2.5 Flash — which is cheaper?

Gemini 2.5 Flash is cheaper on output ($2.5/M vs $3/M).

Which has the larger context window?

Gemini 3 Flash Preview offers the larger context window (1M tokens).

Gemini 3 Flash Preview vs Gemini 2.5 Flash — which should I pick in 2026?

Choose Gemini 2.5 Flash for the lower output price ($2.5/M vs $3/M). Choose Gemini 3 Flash Preview if you need the larger 1M context window.

People also compare

Explore more open-source AI

Browse the full open-source model leaderboard, thousands of tools and live benchmarks — all in one place.

Open the leaderboard →