Gemini 3.6 Flash (Google): live API pricing, context window and the best open-source alternatives, tracked daily by OpenSourceAI.tech.
Prices update automatically — checked hourly against provider list prices.
Independent benchmark scores for Gemini 3.6 Flash, measured by Artificial Analysis. Higher is better. Measurement mode: high.
Gemini 3.6 Flash is a commercial AI model by Google. The specifications below are tracked automatically: pricing is refreshed hourly from public list prices, so the numbers on this page reflect the current cost of using the model through its API.
Gemini 3.6 Flash is an AI language model from Google. It is a proprietary model, available through an API. It scores 50.1 on the Artificial Analysis intelligence index.
Gemini 3.6 Flash is not free: it costs $1.5 per million input tokens and $7.5 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.
Independent benchmarks from Artificial Analysis give it GPQA 92.8%, Humanity's Last Exam 38.3%, Long Context Reasoning 69.7%, SciCode 52.7%, τ-Bench Banking 24.5%, Terminal-Bench 77.5%. It is particularly used for code generation.
It generates about 234.8 tokens per second, with a median 14.56s delay before the first token. Measured independently by Artificial Analysis.
Depending on your workload, an open-weight model that you can self-host may cut costs dramatically or remove per-token pricing entirely. Our comparison pages put this model side by side with the strongest open-source contenders.