Skip to content

Gemini 3.5 Flash Lite

Google · Reasoning · 1M context · released Jul 21, 2026 · checked 9h ago

Ways to buy 8

Way to buyIn / out per 1M
Google Vertex AI
global-flex, cheapest · cached $0.015
$0.15 / $1.25-50% vs list
Google Gemini API
flex · cached $0.015
$0.15 / $1.25-50% vs list
Google Gemini API API
The lab's own API · cached $0.03
$0.3 / $2.5List price
Google Vertex AI
global · cached $0.03
$0.3 / $2.5Same as list
Google Vertex AI
eu · cached $0.033
$0.33 / $2.75+10% vs list
Google Vertex AI
us · cached $0.033
$0.33 / $2.75+10% vs list
Google Vertex AI
global-priority · cached $0.054
$0.54 / $4.5+80% vs list
Google Gemini API
priority · cached $0.054
$0.54 / $4.5+80% vs list

USD per 1M tokens; batch $0.15 in, $1.25 out on the lab's API. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $80.00 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
Gemini 3.8 Flash
Google · 1MCheapest: $0.75 on Google Gemini API (flex)
$0.75 / $3.75
Gemini 3.7 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75100%
Gemini 3.6 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75
Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)
Google · 66K
$0.25 / $1.5
GLM 5.3
Z.ai · 1.31M
$0.563 / $2.5
GLM 4.6
Z.ai · 205K
$0.43 / $1.75
Kimi K2.5
Moonshot AI · 262K
$0.45 / $2.25
Command A+
Cohere · 192K
$0.3 / $1.5

About Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities.

Takes text, image, video, files, audio in and returns text. First listed Jul 21, 2026. Up to 66K output tokens. Blended price $0.85 per 1M tokens at 3 input to 1 output. Model id gemini-3-5-flash-lite.

Price history

$0.3$1$3Jul 21, 2026Output $2.5Input $0.3

List price since Jul 21, 2026, from Internet Archive snapshots of the OpenRouter catalog, then daily reads.

Questions

How much does Gemini 3.5 Flash Lite cost?

$0.3 per million input tokens and $2.5 per million output tokens, with cached input at $0.03, or $0.15 and $1.25 through the Batch API.

What is the context window of Gemini 3.5 Flash Lite?

1M tokens, with up to 66K tokens of output.

Where is Gemini 3.5 Flash Lite cheapest?

Of 8 routes tracked, Google Vertex AI is cheapest at $0.15 input and $1.25 output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.