Qwen3 235B A22B Instruct 2507
Qwen (Alibaba) · Open weights · 262K context · released Jul 21, 2025 · checked 10h ago
Ways to buy 9
| Way to buy | Input | Output | In / out per 1M | Cached | Vs list |
|---|---|---|---|---|---|
| GMICloud fp8, cheapest · cached $0.018 | $0.087 | $0.35 | $0.087 / $0.35-41% vs list | $0.018 | -41% vs list |
fp8 | $0.09 | $0.55 | $0.09 / $0.55-22% vs list | n/a | -22% vs list |
fp8 | $0.09 | $0.58 | $0.09 / $0.58-19% vs list | n/a | -19% vs list |
The lab's own API | $0.15 | $0.598 | $0.15 / $0.598List price | n/a | List price |
fp8 | $0.15 | $0.75 | $0.15 / $0.75+15% vs list | n/a | +15% vs list |
fp8 | $0.2 | $0.6 | $0.2 / $0.6+15% vs list | n/a | +15% vs list |
fp8 · cached $0.05 | $0.14 | $0.8 | $0.14 / $0.8+17% vs list | $0.05 | +17% vs list |
| StreamLake Default region | $0.21 | $0.84 | $0.21 / $0.84+40% vs list | n/a | +40% vs list |
us-south1 | $0.22 | $0.88 | $0.22 / $0.88+47% vs list | n/a | +47% vs list |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $26.91 direct.
Helicone$26.91+0%
Kong AI Gateway$26.91+0%
LiteLLM$26.91+0%
Neon AI Gateway$26.91+0%
Vercel AI Gateway$26.91+0%
Cloudflare AI Gateway$28.26+5%
Requesty$28.26+5%
Eden AI$28.39+5.5%
OpenRouter$28.39+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Qwen (Alibaba) · 1M | $4 | $12 | $4 / $12 | $6 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $0.15 | $0.47 | $0.15 / $0.47 | $0.23 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $2 | $6 | $2 / $6 | $3 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $0.15 | $0.47 | $0.15 / $0.47 | $0.23 | 1M | Lab list price | · |
DeepSeek · 1MCheapest: $0.131 on Morph | $0.15 | $0.6 | $0.15 / $0.6 | $0.262 | 1M | $0.131 on Morph | · |
Google · 1MCheapest: $0.281 on Google Vertex AI (global-flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Vertex AI (global-flex) | · |
Google · 1MCheapest: $0.281 on Google Gemini API (flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Gemini API (flex) | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
About Qwen3 235B A22B Instruct 2507
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass.
Takes text in and returns text. First listed Sep 23, 2026. Up to 236K output tokens. Blended price $0.262 per 1M tokens at 3 input to 1 output. Model id qwen3-235b-a22b-2507.
Price history
List price since Sep 23, 2026, from daily reads of the OpenRouter catalog.
Across fru.devCompany profile
Questions
How much does Qwen3 235B A22B Instruct 2507 cost?
$0.15 per million input tokens and $0.598 per million output tokens.
What is the context window of Qwen3 235B A22B Instruct 2507?
262K tokens, with up to 236K tokens of output.
Where is Qwen3 235B A22B Instruct 2507 cheapest?
Of 9 routes tracked, GMICloud is cheapest at $0.087 input and $0.35 output per million tokens.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.