Ways to buy 2
| Way to buy | Input | Output | In / out per 1M | Cached | Vs list |
|---|---|---|---|---|---|
The lab's own API · cached $0.24 | $1.2 | $4 | $1.2 / $4List price | $0.24 | List price |
fp8 · cached $0.24 | $1.2 | $4 | $1.2 / $4Same as list | $0.24 | Same as list |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $200.00 direct.
Helicone$200.00+0%
Kong AI Gateway$200.00+0%
LiteLLM$200.00+0%
Vercel AI Gateway$200.00+0%
Cloudflare AI Gateway$210.00+5%
Requesty$210.00+5%
Eden AI$211.00+5.5%
OpenRouter$211.00+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Z.ai · 1M | $2.8 | $8.8 | $2.8 / $8.8 | $4.3 | 1M | Lab list price | · |
Z.ai · 1M | $0.37 | $1.25 | $0.37 / $1.25 | $0.59 | 1M | Lab list price | · |
Z.ai · 1.31M | $0.045 | $0.14 | $0.045 / $0.14 | $0.069 | 1.31M | on InferenceNet (fp4) | · |
Z.ai · 1.31M | $0.561 | $1.76 | $0.561 / $1.76 | $0.862 | 1.31M | on Baidu (fp8) | · |
OpenAI · 1.05MCheapest: $2 on OpenAI (flex) | $2 | $10 | $2 / $10 | $4 | 1.05M | $2 on OpenAI (flex) | · |
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
About GLM 5V Turbo
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks.
Takes image, text, video in and returns text. First listed Apr 2, 2026. Up to 131K output tokens. Blended price $1.9 per 1M tokens at 3 input to 1 output. Model id glm-5v-turbo.
Price history
List price since Apr 2, 2026, from Internet Archive snapshots of the OpenRouter catalog, then daily reads.
Across fru.devCompany profile
Questions
How much does GLM 5V Turbo cost?
$1.2 per million input tokens and $4 per million output tokens, with cached input at $0.24.
What is the context window of GLM 5V Turbo?
203K tokens, with up to 131K tokens of output.
Where is GLM 5V Turbo cheapest?
Of 2 routes tracked, Z.ai is cheapest at $1.2 input and $4 output per million tokens.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.