Google · all models ›

GGemini 3.6 FlashAPI

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development.

1MContext window · tokens
$0.75Input price · per M tokens
$3.75Output price · per M tokens
GoogleProvider

Prices update automatically — checked daily against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for Gemini 3.6 Flash, measured by Artificial Analysis. Higher is better. Measurement mode: high.

Intelligence index51.6
Coding index69.2
GPQA92.8%
Humanity's Last Exam40.8%
Long Context Reasoning79%
SciCode52.7%
τ-Bench Banking29.9%
Terminal-Bench77.5%
⚡ Speed186.4 tokens/sec
⏱ Latency12.8s to first token
💰 Blended price$1.5 / 1M tokens
📈 Value34.4 intelligence points per $
vs. models measured heretop 16%
Scores higher than 84% of the 273 models measured by Artificial Analysis and tracked here.
Models at this level cost $1.68 per 1M tokens (median of 14) — this one costs $1.5.
Cheaper and better on this index: GLM 5.3 Flash · Gemini 3.7 Flash · Gemini 3.7 Flash (batch) and 6 more
Benchmark data by Artificial Analysis

About this model

Gemini 3.6 Flash is a commercial AI model by Google. The specifications below are tracked automatically: pricing is refreshed daily from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is Gemini 3.6 Flash?

Gemini 3.6 Flash is an AI language model from Google. It is a proprietary model, available through an API. It scores 51.6 on the Artificial Analysis intelligence index.

Is Gemini 3.6 Flash free?

Gemini 3.6 Flash is not free: it costs $0.75 per million input tokens and $3.75 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is Gemini 3.6 Flash good at?

Independent benchmarks from Artificial Analysis give it GPQA 92.8%, Humanity's Last Exam 40.8%, Long Context Reasoning 79%, SciCode 52.7%, τ-Bench Banking 29.9%, Terminal-Bench 77.5%. It is particularly used for code generation.

How fast is Gemini 3.6 Flash?

It generates about 186.4 tokens per second, with a median 12.8s delay before the first token. Measured independently by Artificial Analysis.

Related models

Gemini 3.1 Pro PreviewGoogleGemini 3.1 Pro Preview Custom ToolsGoogleNano Banana Pro (Gemini 3 Pro Image Preview)GoogleGemini 2.5 ProGoogleGemini 2.5 Pro Preview 05-06GoogleGemini 3.5 FlashGoogle