Ways to buy 14
| Way to buy | Input | Output | In / out per 1M | Cached | Vs cheapest |
|---|---|---|---|---|---|
turbo, fp8 | $0.1 | $0.32 | $0.1 / $0.32Cheapest | n/a | Cheapest |
bf16 | $0.135 | $0.4 | $0.135 / $0.4+30% vs cheapest | n/a | +30% vs cheapest |
| AkashML fp8 · cached $0.1 | $0.2 | $0.52 | $0.2 / $0.52+81% vs cheapest | $0.1 | +81% vs cheapest |
fp8 · cached $0.11 | $0.22 | $0.5 | $0.22 / $0.5+87% vs cheapest | $0.11 | +87% vs cheapest |
Default region | $0.45 | $0.9 | $0.45 / $0.9+263% vs cheapest | n/a | +263% vs cheapest |
Default region · cached $0.295 | $0.59 | $0.79 | $0.59 / $0.79+313% vs cheapest | $0.295 | +313% vs cheapest |
fp16 · cached $0.71 | $0.71 | $0.71 | $0.71 / $0.71+358% vs cheapest | $0.71 | +358% vs cheapest |
Default region | $0.72 | $0.72 | $0.72 / $0.72+365% vs cheapest | n/a | +365% vs cheapest |
us-central1 | $0.72 | $0.72 | $0.72 / $0.72+365% vs cheapest | n/a | +365% vs cheapest |
From the platform pricing page | $0.5 | $1.5 | $0.5 / $1.5+384% vs cheapest | n/a | +384% vs cheapest |
From the platform pricing page | $0.753 | $0.753 | $0.753 / $0.753+386% vs cheapest | n/a | +386% vs cheapest |
fp8 | $0.293 | $2.25 | $0.293 / $2.25+405% vs cheapest | n/a | +405% vs cheapest |
From the platform pricing page | $0.864 | $0.864 | $0.864 / $0.864+457% vs cheapest | n/a | +457% vs cheapest |
Default region | $1.04 | $1.04 | $1.04 / $1.04+571% vs cheapest | n/a | +571% vs cheapest |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $16.40 direct.
Helicone$16.40+0%
Kong AI Gateway$16.40+0%
LiteLLM$16.40+0%
Vercel AI Gateway$16.40+0%
Cloudflare AI Gateway$17.22+5%
Requesty$17.22+5%
Eden AI$17.30+5.5%
OpenRouter$17.30+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
Meta · 1M | $0.1 | $0.2 | $0.1 / $0.2 | $0.125 | 1M | Lab list price | · |
Meta · 1M | $0.1 | $0.2 | $0.1 / $0.2 | $0.125 | 1M | Lab list price | · |
Meta · 131K | $0.3 | $1.1 | $0.3 / $1.1 | $0.5 | 131K | on Phala | · |
DeepSeek · 1MCheapest: $0.131 on Morph | $0.15 | $0.6 | $0.15 / $0.6 | $0.262 | 1M | $0.131 on Morph | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
Qwen (Alibaba) · 1M | $0.15 | $0.47 | $0.15 / $0.47 | $0.23 | 1M | Lab list price | · |
About Llama 3.3 70B Instruct
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out).
Takes text in and returns text. First listed Sep 23, 2026. Up to 16K output tokens. Blended price $0.155 per 1M tokens at 3 input to 1 output. Model id llama-3-3-70b-instruct.
Price history
No list-price history: this model is sold by hosts rather than its lab.
Questions
How much does Llama 3.3 70B Instruct cost?
$0.1 per million input tokens and $0.32 per million output tokens.
What is the context window of Llama 3.3 70B Instruct?
131K tokens, with up to 16K tokens of output.
Where is Llama 3.3 70B Instruct cheapest?
Of 14 routes tracked, DeepInfra is cheapest at $0.1 input and $0.32 output per million tokens.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.