Ways to buy 12
| Way to buy | Input | Output | In / out per 1M | Cached | Vs cheapest |
|---|---|---|---|---|---|
| Darkbloom Default region | $0.042 | $0.22 | $0.042 / $0.22Cheapest | n/a | Cheapest |
| DekaLLM bf16 | $0.06 | $0.33 | $0.06 / $0.33+47% vs cheapest | n/a | +47% vs cheapest |
fp8 | $0.07 | $0.34 | $0.07 / $0.34+59% vs cheapest | n/a | +59% vs cheapest |
| NextBit bf16 · cached $0.05 | $0.09 | $0.3 | $0.09 / $0.3+65% vs cheapest | $0.05 | +65% vs cheapest |
Default region | $0.1 | $0.3 | $0.1 / $0.3+73% vs cheapest | n/a | +73% vs cheapest |
bf16 · cached $0.05 | $0.1 | $0.3 | $0.1 / $0.3+73% vs cheapest | $0.05 | +73% vs cheapest |
| Makora Default region · cached $0.034 | $0.1 | $0.34 | $0.1 / $0.34+85% vs cheapest | $0.034 | +85% vs cheapest |
bf16 | $0.13 | $0.4 | $0.13 / $0.4+128% vs cheapest | n/a | +128% vs cheapest |
bf16 · cached $0.05 | $0.13 | $0.4 | $0.13 / $0.4+128% vs cheapest | $0.05 | +128% vs cheapest |
bf16 · cached $0.05 | $0.13 | $0.4 | $0.13 / $0.4+128% vs cheapest | $0.05 | +128% vs cheapest |
fp8 · cached $0.05 | $0.14 | $0.4 | $0.14 / $0.4+137% vs cheapest | $0.05 | +137% vs cheapest |
global | $0.15 | $0.6 | $0.15 / $0.6+203% vs cheapest | n/a | +203% vs cheapest |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $8.60 direct.
Helicone$8.60+0%
Kong AI Gateway$8.60+0%
LiteLLM$8.60+0%
Vercel AI Gateway$8.60+0%
Cloudflare AI Gateway$9.03+5%
Requesty$9.03+5%
Eden AI$9.07+5.5%
OpenRouter$9.07+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75100% | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | 100% |
Google · 1MCheapest: $0.425 on Google Vertex AI (global-flex) | $0.3 | $2.5 | $0.3 / $2.5 | $0.85 | 1M | $0.425 on Google Vertex AI (global-flex) | · |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | · |
Meta · 1M | $0.1 | $0.2 | $0.1 / $0.2 | $0.125 | 1M | Lab list price | · |
Z.ai · 1.31M | $0.045 | $0.14 | $0.045 / $0.14 | $0.069 | 1.31M | on InferenceNet (fp4) | · |
Meta · 1M | $0.1 | $0.2 | $0.1 / $0.2 | $0.125 | 1M | Lab list price | · |
DeepSeek · 1.31M | $0.053 | $0.158 | $0.053 / $0.158 | $0.079 | 1.31M | on StreamLake (fp8) | · |
About Gemma 4 26B A4B
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind.
Takes image, text, video in and returns text. First listed Sep 23, 2026. Up to 236K output tokens. Blended price $0.086 per 1M tokens at 3 input to 1 output. Model id gemma-4-26b-a4b-it.
Price history
No list-price history: this model is sold by hosts rather than its lab.
Across fru.devCompany profile7 acquisitionsPay schedule7 conferences
Questions
How much does Gemma 4 26B A4B cost?
$0.042 per million input tokens and $0.22 per million output tokens.
What is the context window of Gemma 4 26B A4B?
262K tokens, with up to 236K tokens of output.
Where is Gemma 4 26B A4B cheapest?
Of 12 routes tracked, Darkbloom is cheapest at $0.042 input and $0.22 output per million tokens.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.