Qwen2.5 Coder 32B Instruct
Qwen (Alibaba) · Open weights · 33K context · released Nov 11, 2024 · checked 6h ago
Ways to buy 1
| Way to buy | Input | Output | In / out per 1M | Cached | Vs cheapest |
|---|---|---|---|---|---|
Default region | $0.66 | $1 | $0.66 / $1Cheapest | n/a | Cheapest |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $86.00 direct.
Helicone$86.00+0%
Kong AI Gateway$86.00+0%
LiteLLM$86.00+0%
Vercel AI Gateway$86.00+0%
Cloudflare AI Gateway$90.30+5%
Requesty$90.30+5%
Eden AI$90.73+5.5%
OpenRouter$90.73+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Qwen (Alibaba) · 1M | $4 | $12 | $4 / $12 | $6 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $0.15 | $0.47 | $0.15 / $0.47 | $0.23 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $2 | $6 | $2 / $6 | $3 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $0.15 | $0.47 | $0.15 / $0.47 | $0.23 | 1M | Lab list price | · |
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Z.ai · 1.31M | $0.561 | $1.76 | $0.561 / $1.76 | $0.862 | 1.31M | on Baidu (fp8) | · |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75100% | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | 100% |
Z.ai · 205K | $0.43 | $1.75 | $0.43 / $1.75 | $0.76 | 205K | on Venice (fp4) | · |
About Qwen2.5 Coder 32B Instruct
Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen).
Takes text in and returns text. First listed Sep 23, 2026. Up to 29K output tokens. Blended price $0.745 per 1M tokens at 3 input to 1 output. Model id qwen-2-5-coder-32b-instruct.
Price history
No list-price history: this model is sold by hosts rather than its lab.
Across fru.devCompany profile
Questions
How much does Qwen2.5 Coder 32B Instruct cost?
$0.66 per million input tokens and $1 per million output tokens.
What is the context window of Qwen2.5 Coder 32B Instruct?
33K tokens, with up to 29K tokens of output.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.