Skip to content

Qwen2.5 VL 72B Instruct

Qwen (Alibaba) · Open weights · 128K context · released Feb 1, 2025 · checked 8h ago

Ways to buy 1

Way to buyIn / out per 1M
Parasail
fp8 · cached $0.4
$0.8 / $1Cheapest

USD per 1M tokens. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $100.00 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
Qwen3.8 Max Prime
Qwen (Alibaba) · 1M
$4 / $12
Qwen3.8 Omni Flash
Qwen (Alibaba) · 1M
$0.15 / $0.47
Qwen3.8 Max (0902)
Qwen (Alibaba) · 1M
$2 / $6
Qwen3.8 Flash
Qwen (Alibaba) · 1M
$0.15 / $0.47
Gemini 3.8 Flash
Google · 1MCheapest: $0.75 on Google Gemini API (flex)
$0.75 / $3.75
GLM 5.3
Z.ai · 1.31M
$0.561 / $1.76
Gemini 3.7 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75100%
GLM 4.6
Z.ai · 205K
$0.43 / $1.75

About Qwen2.5 VL 72B Instruct

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects.

Takes text, image in and returns text. First listed Sep 23, 2026. Up to 115K output tokens. Blended price $0.85 per 1M tokens at 3 input to 1 output. Model id qwen2-5-vl-72b-instruct.

Price history

No list-price history: this model is sold by hosts rather than its lab.

Across fru.devCompany profile

Questions

How much does Qwen2.5 VL 72B Instruct cost?

$0.8 per million input tokens and $1 per million output tokens, with cached input at $0.4.

What is the context window of Qwen2.5 VL 72B Instruct?

128K tokens, with up to 115K tokens of output.

Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.