Ways to buy 2
| Way to buy | Input | Output | In / out per 1M | Cached | Vs list |
|---|---|---|---|---|---|
The lab's own API | $3 | $4 | $3 / $4List price | n/a | List price |
Default region | $3 | $4 | $3 / $4Same as list | n/a | Same as list |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $380.00 direct.
Helicone$380.00+0%
Kong AI Gateway$380.00+0%
LiteLLM$380.00+0%
Vercel AI Gateway$380.00+0%
Cloudflare AI Gateway$399.00+5%
Requesty$399.00+5%
Eden AI$400.90+5.5%
OpenRouter$400.90+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $2 on OpenAI (flex) | $2 | $10 | $2 / $10 | $4 | 1.05M | $2 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $2 on OpenAI (flex) | $2 | $10 | $2 / $10 | $4 | 1.05M | $2 on OpenAI (flex) | · |
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Google · 1MCheapest: $2.25 on Google Vertex AI (global-flex) | $2 | $12 | $2 / $12 | $4.5 | 1M | $2.25 on Google Vertex AI (global-flex) | · |
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
About GPT-3.5 Turbo 16k
This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost.
Takes text in and returns text. First listed Mar 26, 2024. Up to 4K output tokens. Blended price $3.25 per 1M tokens at 3 input to 1 output. Model id gpt-3-5-turbo-16k.
Price history
List price since Mar 26, 2024, from Internet Archive snapshots of the OpenRouter catalog, then daily reads.
Questions
How much does GPT-3.5 Turbo 16k cost?
$3 per million input tokens and $4 per million output tokens.
What is the context window of GPT-3.5 Turbo 16k?
16K tokens, with up to 4K tokens of output.
Where is GPT-3.5 Turbo 16k cheapest?
Of 2 routes tracked, Azure AI Foundry is cheapest at $3 input and $4 output per million tokens.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.