Skip to content

Gemini 3.1 Flash Lite

Google · Reasoning · 1M context · released May 7, 2026 · checked 6h ago

Ways to buy 8

Way to buyIn / out per 1M
Google Vertex AI
global-flex, cheapest · cached $0.013
$0.125 / $0.75-50% vs list
Google Gemini API
flex · cached $0.013
$0.125 / $0.75-50% vs list
Google Gemini API API
The lab's own API · cached $0.025
$0.25 / $1.5List price
Google Vertex AI
global · cached $0.025
$0.25 / $1.5Same as list
Google Vertex AI
eu · cached $0.028
$0.275 / $1.65+10% vs list
Google Vertex AI
us · cached $0.028
$0.275 / $1.65+10% vs list
Google Vertex AI
global-priority · cached $0.045
$0.45 / $2.7+80% vs list
Google Gemini API
priority · cached $0.045
$0.45 / $2.7+80% vs list

USD per 1M tokens; batch $0.125 in, $0.75 out on the lab's API. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $55.00 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
Gemini 3.8 Flash
Google · 1MCheapest: $0.75 on Google Gemini API (flex)
$0.75 / $3.75
Gemini 3.7 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75100%
Gemini 3.5 Flash Lite
Google · 1MCheapest: $0.425 on Google Vertex AI (global-flex)
$0.3 / $2.5
Gemini 3.6 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75
GLM 5.3
Z.ai · 1.31M
$0.561 / $1.76
GLM 4.6
Z.ai · 205K
$0.43 / $1.75
DeepSeek V4.1 Flash
DeepSeek · 1MCheapest: $0.131 on Morph
$0.15 / $0.6
Kimi K2.5
Moonshot AI · 262K
$0.45 / $2.25

About Gemini 3.1 Flash Lite

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.

Takes text, image, video, files, audio in and returns text. First listed May 13, 2026. Up to 66K output tokens. Blended price $0.563 per 1M tokens at 3 input to 1 output. Overall score 1 on leadersboard.fru.dev (rank 59). Model id gemini-3-1-flash-lite.

Price history

$0.3$1May 13, 2026Output $1.5Input $0.25

List price since May 13, 2026, from Internet Archive snapshots of the OpenRouter catalog, then daily reads.

Questions

How much does Gemini 3.1 Flash Lite cost?

$0.25 per million input tokens and $1.5 per million output tokens, with cached input at $0.025, or $0.125 and $0.75 through the Batch API.

What is the context window of Gemini 3.1 Flash Lite?

1M tokens, with up to 66K tokens of output.

Where is Gemini 3.1 Flash Lite cheapest?

Of 8 routes tracked, Google Vertex AI is cheapest at $0.125 input and $0.75 output per million tokens.

Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.