Google

GGemini 3.5 FlashAPI

Gemini 3.5 Flash (Google): live API pricing, context window and the best open-source alternatives, tracked daily by OpenSourceAI.tech.

Context window
1M
tokens
Input price
$1.5
per M tokens
Output price
$9
per M tokens
Provider
Google

Prices update automatically — checked hourly against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for Gemini 3.5 Flash, measured by Artificial Analysis. Higher is better. Measurement mode: high.

Intelligence index50.2
Coding index70.1
GPQA92.2%
Humanity's Last Exam41%
Long Context Reasoning69.3%
SciCode53.1%
IFBench76.3%
τ²-Bench95.3%
τ-Bench Banking25.4%
Terminal-Bench78.7%
Terminal-Bench Hard40.9%
⚡ Speed253.5 tokens/sec
⏱ Latency15.52s to first token
💰 Blended price$3.375 / 1M tokens
📈 Value14.9 intelligence points per $
Benchmark data by Artificial Analysis

About this model

Gemini 3.5 Flash is a commercial AI model by Google. The specifications below are tracked automatically: pricing is refreshed hourly from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is Gemini 3.5 Flash?

Gemini 3.5 Flash is an AI language model from Google. It is a proprietary model, available through an API. It scores 50.2 on the Artificial Analysis intelligence index.

Is Gemini 3.5 Flash free?

Gemini 3.5 Flash is not free: it costs $1.5 per million input tokens and $9 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is Gemini 3.5 Flash good at?

Independent benchmarks from Artificial Analysis give it GPQA 92.2%, Humanity's Last Exam 41%, Long Context Reasoning 69.3%, SciCode 53.1%, IFBench 76.3%, τ²-Bench 95.3%, τ-Bench Banking 25.4%, Terminal-Bench 78.7%, Terminal-Bench Hard 40.9%. It is particularly used for code generation.

How fast is Gemini 3.5 Flash?

It generates about 253.5 tokens per second, with a median 15.52s delay before the first token. Measured independently by Artificial Analysis.

Open-source alternatives

Depending on your workload, an open-weight model that you can self-host may cut costs dramatically or remove per-token pricing entirely. Our comparison pages put this model side by side with the strongest open-source contenders.

Gemini 3.1 Pro PreviewGoogleGemini 2.5 ProGoogleGemini 3.6 FlashGoogleGemini 3 Flash PreviewGoogleGemini 2.5 FlashGoogleGemini 3.5 Flash LiteGoogle