Skip to content

Qwen3.5 397B A17B

Qwen (Alibaba) · Open weights · 262K context · released Feb 16, 2026 · checked 8h ago

Ways to buy 10

Way to buyIn / out per 1M
Qwen (Alibaba) API
The lab's own API
$0.39 / $2.34List price
DeepInfra
fp8 · cached $0.22
$0.45 / $3+24% vs list
Parasail
fp8 · cached $0.3
$0.5 / $3.6+45% vs list
AtlasCloud
fp8 · cached $0.55
$0.55 / $3.5+47% vs list
DigitalOcean
Default region · cached $0.11
$0.55 / $3.5+47% vs list
Phala
Default region · cached $0.225
$0.55 / $3.5+47% vs list
GMICloud
fp8
$0.6 / $3.6+54% vs list
Novita AI
Default region
$0.6 / $3.6+54% vs list
StreamLake
Default region · cached $0.12
$0.6 / $3.6+54% vs list
Venice
Default region
$0.75 / $4.5+92% vs list

USD per 1M tokens. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $85.80 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
Qwen3.8 Max Prime
Qwen (Alibaba) · 1M
$4 / $12
Qwen3.8 Omni Flash
Qwen (Alibaba) · 1M
$0.15 / $0.47
Qwen3.8 Max (0902)
Qwen (Alibaba) · 1M
$2 / $6
Qwen3.8 Flash
Qwen (Alibaba) · 1M
$0.15 / $0.47
Gemini 3.8 Flash
Google · 1MCheapest: $0.75 on Google Gemini API (flex)
$0.75 / $3.75
GLM 5.3
Z.ai · 1.31M
$0.561 / $1.76
Gemini 3.7 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75100%
GLM 4.6
Z.ai · 205K
$0.43 / $1.75

About Qwen3.5 397B A17B

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency.

Takes text, image, video in and returns text. First listed Sep 23, 2026. Up to 236K output tokens. Blended price $0.878 per 1M tokens at 3 input to 1 output. Model id qwen3-5-397b-a17b.

Price history

$0.3$1$3Sep 23, 2026Output $2.34Input $0.39

List price since Sep 23, 2026, from daily reads of the OpenRouter catalog.

Across fru.devCompany profile

Questions

How much does Qwen3.5 397B A17B cost?

$0.39 per million input tokens and $2.34 per million output tokens.

What is the context window of Qwen3.5 397B A17B?

262K tokens, with up to 236K tokens of output.

Where is Qwen3.5 397B A17B cheapest?

Of 10 routes tracked, Qwen (Alibaba) is cheapest at $0.39 input and $2.34 output per million tokens.

Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.