Google · all models ›

GGemini 3.5 Flash LiteAPI

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities.

1MContext window · tokens
$0.3Input price · per M tokens
$2.5Output price · per M tokens
GoogleProvider

Prices update automatically — checked daily against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for Gemini 3.5 Flash Lite, measured by Artificial Analysis. Higher is better.

Intelligence index37.4
Coding index49.3
GPQA83.8%
Humanity's Last Exam18.8%
Long Context Reasoning74.7%
SciCode40.9%
τ-Bench Banking17.5%
Terminal-Bench53.6%
⚡ Speed366.6 tokens/sec
⏱ Latency7.14s to first token
💰 Blended price$0.85 / 1M tokens
📈 Value44 intelligence points per $
vs. models measured heretop 35%
Scores higher than 65% of the 273 models measured by Artificial Analysis and tracked here.
Models at this level cost $1.72 per 1M tokens (median of 25) — this one costs $0.85.
Cheaper and better on this index: GLM 5.3 Flash · Gemini 3.7 Flash · Gemini 3.7 Flash (batch) and 21 more
Benchmark data by Artificial Analysis

About this model

Gemini 3.5 Flash Lite is a commercial AI model by Google. The specifications below are tracked automatically: pricing is refreshed daily from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is Gemini 3.5 Flash Lite?

Gemini 3.5 Flash Lite is an AI language model from Google. It is a proprietary model, available through an API. It scores 37.4 on the Artificial Analysis intelligence index.

Is Gemini 3.5 Flash Lite free?

Gemini 3.5 Flash Lite is not free: it costs $0.3 per million input tokens and $2.5 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is Gemini 3.5 Flash Lite good at?

Independent benchmarks from Artificial Analysis give it GPQA 83.8%, Humanity's Last Exam 18.8%, Long Context Reasoning 74.7%, SciCode 40.9%, τ-Bench Banking 17.5%, Terminal-Bench 53.6%. It is particularly used for code generation.

How fast is Gemini 3.5 Flash Lite?

It generates about 366.6 tokens per second, with a median 7.14s delay before the first token. Measured independently by Artificial Analysis.

Related models

Gemini 3.1 Pro PreviewGoogleGemini 3.1 Pro Preview Custom ToolsGoogleNano Banana Pro (Gemini 3 Pro Image Preview)GoogleGemini 2.5 ProGoogleGemini 2.5 Pro Preview 05-06GoogleGemini 3.5 FlashGoogle