Skip to content

Qwen3 235B A22B Instruct 2507

Qwen (Alibaba) · Open weights · 262K context · released Jul 21, 2025 · checked 10h ago

Ways to buy 9

Way to buyIn / out per 1M
GMICloud
fp8, cheapest · cached $0.018
$0.087 / $0.35-41% vs list
DeepInfra
fp8
$0.09 / $0.55-22% vs list
Novita AI
fp8
$0.09 / $0.58-19% vs list
Qwen (Alibaba) API
The lab's own API
$0.15 / $0.598List price
Venice
fp8
$0.15 / $0.75+15% vs list
Nebius
fp8
$0.2 / $0.6+15% vs list
Parasail
fp8 · cached $0.05
$0.14 / $0.8+17% vs list
StreamLake
Default region
$0.21 / $0.84+40% vs list
Google Vertex AI
us-south1
$0.22 / $0.88+47% vs list

USD per 1M tokens. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $26.91 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
Qwen3.8 Max Prime
Qwen (Alibaba) · 1M
$4 / $12
Qwen3.8 Omni Flash
Qwen (Alibaba) · 1M
$0.15 / $0.47
Qwen3.8 Max (0902)
Qwen (Alibaba) · 1M
$2 / $6
Qwen3.8 Flash
Qwen (Alibaba) · 1M
$0.15 / $0.47
DeepSeek V4.1 Flash
DeepSeek · 1MCheapest: $0.131 on Morph
$0.15 / $0.6
Gemini 3.1 Flash Lite
Google · 1MCheapest: $0.281 on Google Vertex AI (global-flex)
$0.25 / $1.5
Gemini 3.1 Flash Lite Preview
Google · 1MCheapest: $0.281 on Google Gemini API (flex)
$0.25 / $1.5
GPT-6 Luna
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex)
$0.1 / $0.5

About Qwen3 235B A22B Instruct 2507

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass.

Takes text in and returns text. First listed Sep 23, 2026. Up to 236K output tokens. Blended price $0.262 per 1M tokens at 3 input to 1 output. Model id qwen3-235b-a22b-2507.

Price history

$0.3Sep 23, 2026Output $0.598Input $0.15

List price since Sep 23, 2026, from daily reads of the OpenRouter catalog.

Across fru.devCompany profile

Questions

How much does Qwen3 235B A22B Instruct 2507 cost?

$0.15 per million input tokens and $0.598 per million output tokens.

What is the context window of Qwen3 235B A22B Instruct 2507?

262K tokens, with up to 236K tokens of output.

Where is Qwen3 235B A22B Instruct 2507 cheapest?

Of 9 routes tracked, GMICloud is cheapest at $0.087 input and $0.35 output per million tokens.

Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.