Ways to buy 2
| Way to buy | Input | Output | In / out per 1M | Cached | Vs list |
|---|---|---|---|---|---|
The lab's own API | $1 | $2 | $1 / $2List price | n/a | List price |
Default region | $1 | $2 | $1 / $2Same as list | n/a | Same as list |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $140.00 direct.
Helicone$140.00+0%
Kong AI Gateway$140.00+0%
LiteLLM$140.00+0%
Vercel AI Gateway$140.00+0%
Cloudflare AI Gateway$147.00+5%
Requesty$147.00+5%
Eden AI$147.70+5.5%
OpenRouter$147.70+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $2 on OpenAI (flex) | $2 | $10 | $2 / $10 | $4 | 1.05M | $2 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $2 on OpenAI (flex) | $2 | $10 | $2 / $10 | $4 | 1.05M | $2 on OpenAI (flex) | · |
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
Z.ai · 1.31M | $0.561 | $1.76 | $0.561 / $1.76 | $0.862 | 1.31M | on Baidu (fp8) | · |
About GPT-3.5 Turbo (older v0613)
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks.
Takes text in and returns text. First listed Mar 26, 2024. Up to 4K output tokens. Blended price $1.25 per 1M tokens at 3 input to 1 output. Model id gpt-3-5-turbo-0613.
Price history
List price since Mar 26, 2024, from Internet Archive snapshots of the OpenRouter catalog, then daily reads.
Questions
How much does GPT-3.5 Turbo (older v0613) cost?
$1 per million input tokens and $2 per million output tokens.
What is the context window of GPT-3.5 Turbo (older v0613)?
4K tokens, with up to 4K tokens of output.
Where is GPT-3.5 Turbo (older v0613) cheapest?
Of 2 routes tracked, Azure AI Foundry is cheapest at $1 input and $2 output per million tokens.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.