Ways to buy 2
| Way to buy | Input | Output | In / out per 1M | Cached | Vs cheapest |
|---|---|---|---|---|---|
bf16 · cached $0.055 | $0.3 | $0.9 | $0.3 / $0.9Cheapest | $0.055 | Cheapest |
fp8 · cached $0.05 | $0.3 | $0.9 | $0.3 / $0.9Cheapest | $0.05 | Cheapest |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $48.00 direct.
Helicone$48.00+0%
Kong AI Gateway$48.00+0%
LiteLLM$48.00+0%
Vercel AI Gateway$48.00+0%
Cloudflare AI Gateway$50.40+5%
Requesty$50.40+5%
Eden AI$50.64+5.5%
OpenRouter$50.64+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Z.ai · 1M | $2.8 | $8.8 | $2.8 / $8.8 | $4.3 | 1M | Lab list price | · |
Z.ai · 1M | $0.37 | $1.25 | $0.37 / $1.25 | $0.59 | 1M | Lab list price | · |
Z.ai · 1.31M | $0.045 | $0.14 | $0.045 / $0.14 | $0.069 | 1.31M | on InferenceNet (fp4) | · |
Z.ai · 1.31M | $0.561 | $1.76 | $0.561 / $1.76 | $0.862 | 1.31M | on Baidu (fp8) | · |
DeepSeek · 1MCheapest: $0.131 on Morph | $0.15 | $0.6 | $0.15 / $0.6 | $0.262 | 1M | $0.131 on Morph | · |
Moonshot AI · 262K | $0.45 | $2.25 | $0.45 / $2.25 | $0.9 | 262K | on SiliconFlow (int4) | · |
Google · 1MCheapest: $0.281 on Google Vertex AI (global-flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Vertex AI (global-flex) | · |
Google · 1MCheapest: $0.281 on Google Gemini API (flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Gemini API (flex) | · |
About GLM 4.6V
GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media.
Takes image, text, video in and returns text. First listed Sep 23, 2026. Up to 33K output tokens. Blended price $0.45 per 1M tokens at 3 input to 1 output. Model id glm-4-6v.
Price history
No list-price history: this model is sold by hosts rather than its lab.
Across fru.devCompany profile
Questions
How much does GLM 4.6V cost?
$0.3 per million input tokens and $0.9 per million output tokens, with cached input at $0.055.
What is the context window of GLM 4.6V?
131K tokens, with up to 33K tokens of output.
Where is GLM 4.6V cheapest?
Of 2 routes tracked, Novita AI is cheapest at $0.3 input and $0.9 output per million tokens.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.