Skip to content

Compare model prices

Input, output, cached and batch prices, every route and a month of tokens, side by side.

Gemini Embedding 001Gemini 3.8 Flash

A month of tokens

  • Gemini Embedding 001$15.00
  • Gemini 3.8 Flash$150.00

100M input and 20M output tokens, list price. Use your own numbers.

Side by side

Gemini Embedding 001Gemini 3.8 Flash
Input per 1M$0.15$0.75
Output per 1MFree$3.75
Cached inputn/a$0.075
Batch in / outn/a$0.375 / $1.88
Blended (3:1)$0.112$1.5
Context20K1M
Routes26
ReleasedOct 31, 2025Sep 2, 2026
Leadersboard scoren/a15 (#13)

On the price map

Size: context window, 8K to 2MEqual blended cost (3:1) 245 models, lower left is cheaper

Every route

Gemini Embedding 001

Way to buyIn / out per 1M
Google Gemini API API
The lab's own API
$0.15 / FreeList price
Google Vertex AI
us-central1
$0.15 / FreeSame as list

Gemini 3.8 Flash

Way to buyIn / out per 1M
Google Vertex AI
global-flex, cheapest · cached $0.037
$0.375 / $1.88-50% vs list
Google Gemini API
flex · cached $0.037
$0.375 / $1.88-50% vs list
Google Gemini API API
The lab's own API · cached $0.075
$0.75 / $3.75List price
Google Vertex AI
global · cached $0.075
$0.75 / $3.75Same as list
Google Vertex AI
global-priority · cached $0.135
$1.35 / $6.75+80% vs list
Google Gemini API
priority · cached $0.135
$1.35 / $6.75+80% vs list

Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.