Skip to content

Gemini 3.1 Flash Lite

API pricing per million tokens

gemini-3-1-flash-lite ReasoningChecked 8h ago
Input
$0.25
per 1M, list
Output
$1.5
per 1M tokens
Cached input
$0.025
batch $0.125 / $0.75
Context
1M
66K max output

Where to buy it (8)

  • Google Vertex AIglobal-flex-50%$0.125 / $0.75cached $0.013
  • Google Gemini APIflex-50%$0.125 / $0.75cached $0.013
  • Google Vertex AIglobal$0.25 / $1.5cached $0.025
  • Google Gemini API (list)the lab's own API$0.25 / $1.5cached $0.025
  • Google Vertex AIeu+10%$0.275 / $1.65cached $0.028
  • Google Vertex AIus+10%$0.275 / $1.65cached $0.028
  • Google Vertex AIglobal-priority+80%$0.45 / $2.7cached $0.045
  • Google Gemini APIpriority+80%$0.45 / $2.7cached $0.045

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $55.00 direct.

Related models

About Gemini 3.1 Flash Lite

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.

Takes text, image, video, files, audio in and returns text. First listed May 13, 2026. Blended price $0.563 per 1M tokens at 3 input to 1 output. Overall score 1 on leadersboard.fru.dev (rank 60).

Questions

How much does Gemini 3.1 Flash Lite cost?

$0.25 per million input tokens and $1.5 per million output tokens, with cached input at $0.025, or $0.125 and $0.75 through the Batch API.

What is the context window of Gemini 3.1 Flash Lite?

1M tokens, with up to 66K tokens of output.

Where is Gemini 3.1 Flash Lite cheapest?

Of 8 routes tracked, Google Vertex AI is cheapest at $0.125 input and $0.75 output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.