Gemini 3.1 Flash Lite vs Gemini 2.5 Flash Lite compared — price per token, context window, multimodality, openness and which to choose. Full 2026 breakdown.
Gemini 3.1 Flash Lite — full profile › · Gemini 2.5 Flash Lite — full profile ›
Prices & specs refreshed from live data · OpenSourceAI.tech
Open-model prices = cheapest provider via OpenRouter; official maker rates may be higher.
| Spec | Gemini 3.1 Flash Lite | Gemini 2.5 Flash Lite |
|---|---|---|
| Maker | ||
| Type | Proprietary | Proprietary |
| Context window | 1M tokens | 1M tokens |
| Input price | $0.25/M | $0.1/M |
| Output price | $1.5/M | $0.4/M |
| Vision / multimodal | Yes | Yes |
| Tool / function calling | Yes | Yes |
| Self-hostable | No (API only) | No (API only) |
| License | Proprietary | Proprietary |
| Capability | Gemini 3.1 Flash Lite | Gemini 2.5 Flash Lite |
|---|---|---|
| Open weights (downloadable) | ✗ | ✗ |
| Self-hostable | ✗ | ✗ |
| Runs fully offline | ✗ | ✗ |
| Vision / multimodal | ✓ | ✓ |
| Tool / function calling | ✓ | ✓ |
| 1M+ context window | ✓ | ✓ |
Independent benchmark scores measured by Artificial Analysis. Higher is better (except latency).
Benchmark data by Artificial Analysis.
| Criterion | Gemini 3.1 Flash Lite | Gemini 2.5 Flash Lite |
|---|---|---|
| Cost-efficiency | 4.5 | 5.0 |
| Context window | 5.0 | 5.0 |
| Openness | 1.5 | 1.5 |
| Self-hosting | 1.0 | 1.0 |
| Multimodality | 5.0 | 5.0 |
Scores come from live data — output price (cost), context length, open vs closed weights (openness & self-hosting) and vision/tool support (multimodality). Raw task quality isn't scored here; it depends on your benchmark — see the verdict.
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
These variants are tracked but not compared here — one page per family keeps the comparison readable.
Gemini 2.5 Flash Lite is cheaper on output ($0.4/M vs $1.5/M).
Gemini 3.1 Flash Lite offers the larger context window (1M tokens).
Choose Gemini 2.5 Flash Lite for the lower output price ($0.4/M vs $1.5/M). Choose Gemini 3.1 Flash Lite if you need the larger 1M context window.
Browse the full open-source model leaderboard, thousands of tools and live benchmarks — all in one place.
Open the leaderboard →