Qwen3 VL 8B Thinking
Qwen (Alibaba) · Open weights · 131K context · released Oct 14, 2025 · checked 7h ago
Ways to buy 1
| Way to buy | Input | Output | In / out per 1M | Cached | Vs list |
|---|---|---|---|---|---|
The lab's own API | $0.18 | $2.1 | $0.18 / $2.1List price | n/a | List price |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $60.00 direct.
Helicone$60.00+0%
Kong AI Gateway$60.00+0%
LiteLLM$60.00+0%
Vercel AI Gateway$60.00+0%
Cloudflare AI Gateway$63.00+5%
Requesty$63.00+5%
Eden AI$63.30+5.5%
OpenRouter$63.30+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Qwen (Alibaba) · 1M | $4 | $12 | $4 / $12 | $6 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $0.15 | $0.47 | $0.15 / $0.47 | $0.23 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $2 | $6 | $2 / $6 | $3 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $0.15 | $0.47 | $0.15 / $0.47 | $0.23 | 1M | Lab list price | · |
Z.ai · 1.31M | $0.561 | $1.76 | $0.561 / $1.76 | $0.862 | 1.31M | on Baidu (fp8) | · |
Z.ai · 205K | $0.43 | $1.75 | $0.43 / $1.75 | $0.76 | 205K | on Venice (fp4) | · |
Moonshot AI · 262K | $0.45 | $2.25 | $0.45 / $2.25 | $0.9 | 262K | on SiliconFlow (int4) | · |
Google · 1MCheapest: $0.281 on Google Vertex AI (global-flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Vertex AI (global-flex) | · |
About Qwen3 VL 8B Thinking
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences.
Takes image, text in and returns text. First listed Sep 23, 2026. Up to 33K output tokens. Blended price $0.66 per 1M tokens at 3 input to 1 output. Model id qwen3-vl-8b-thinking.
Price history
List price since Sep 23, 2026, from daily reads of the OpenRouter catalog.
Across fru.devCompany profile
Questions
How much does Qwen3 VL 8B Thinking cost?
$0.18 per million input tokens and $2.1 per million output tokens.
What is the context window of Qwen3 VL 8B Thinking?
131K tokens, with up to 33K tokens of output.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.