Skip to content

Gemini 3.5 Flash

Google · Reasoning · 1M context · released May 19, 2026 · checked 47h ago

Ways to buy 9

Way to buyIn / out per 1M
Google Vertex AI
global-flex, cheapest · cached $0.075
$0.75 / $4.5-50% vs list
Google Gemini API
flex · cached $0.075
$0.75 / $4.5-50% vs list
Google Gemini API API
The lab's own API · cached $0.15
$1.5 / $9List price
Google Vertex AI
global · cached $0.15
$1.5 / $9Same as list
Google Vertex AI
us · cached $0.165
$1.65 / $9.9+10% vs list
Snowflake Cortex AI (AI_COMPLETE)
From the platform pricing page
$1.8 / $10.8+20% vs list
Databricks Mosaic AI Gateway + Foundation Model APIs
From the platform pricing page
$1.88 / $11.3+25% vs list
Google Vertex AI
global-priority · cached $0.27
$2.7 / $16.2+80% vs list
Google Gemini API
priority · cached $0.27
$2.7 / $16.2+80% vs list

USD per 1M tokens; batch $0.75 in, $4.5 out on the lab's API. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $330.00 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
Gemini 3.8 Flash
Google · 1MCheapest: $0.75 on Google Gemini API (flex)
$0.75 / $3.75
Gemini 3.7 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75100%
Gemini 3.5 Flash Lite
Google · 1MCheapest: $0.425 on Google Vertex AI (global-flex)
$0.3 / $2.5
Gemini 3.6 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75
GPT-5.6 Sol
OpenAI · 1.05MCheapest: $2 on OpenAI (flex)
$2 / $10
Muse Spark 1.3
Meta · 1M
$1.25 / $4.25
Muse Spark 1.1
Meta · 1M
$1.25 / $4.25
Qwen3.8 Max (0902)
Qwen (Alibaba) · 1M
$2 / $6

About Gemini 3.5 Flash

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed.

Takes text, image, video, files, audio in and returns text. First listed May 20, 2026. Up to 66K output tokens. Blended price $3.38 per 1M tokens at 3 input to 1 output. Overall score 3 on leadersboard.fru.dev (rank 35). Model id gemini-3-5-flash.

Price history

$3$10May 20, 2026Output $9Input $1.5

List price since May 20, 2026, from Internet Archive snapshots of the OpenRouter catalog, then daily reads.

Questions

How much does Gemini 3.5 Flash cost?

$1.5 per million input tokens and $9 per million output tokens, with cached input at $0.15, or $0.75 and $4.5 through the Batch API.

What is the context window of Gemini 3.5 Flash?

1M tokens, with up to 66K tokens of output.

Where is Gemini 3.5 Flash cheapest?

Of 9 routes tracked, Google Vertex AI is cheapest at $0.75 input and $4.5 output per million tokens.

Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.