Ways to buy 4
| Way to buy | Input | Output | In / out per 1M | Cached | Vs cheapest |
|---|---|---|---|---|---|
fp4 · cached $0.08 | $0.43 | $1.75 | $0.43 / $1.75Cheapest | $0.08 | Cheapest |
fp4 · cached $0.1 | $0.5 | $2 | $0.5 / $2+15% vs cheapest | $0.1 | +15% vs cheapest |
bf16 · cached $0.11 | $0.55 | $2.2 | $0.55 / $2.2+27% vs cheapest | $0.11 | +27% vs cheapest |
fp4 · cached $0.11 | $0.6 | $2.2 | $0.6 / $2.2+32% vs cheapest | $0.11 | +32% vs cheapest |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $78.00 direct.
Helicone$78.00+0%
Kong AI Gateway$78.00+0%
LiteLLM$78.00+0%
Vercel AI Gateway$78.00+0%
Cloudflare AI Gateway$81.90+5%
Requesty$81.90+5%
Eden AI$82.29+5.5%
OpenRouter$82.29+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Z.ai · 1M | $2.8 | $8.8 | $2.8 / $8.8 | $4.3 | 1M | Lab list price | · |
Z.ai · 1M | $0.37 | $1.25 | $0.37 / $1.25 | $0.59 | 1M | Lab list price | · |
Z.ai · 1.31M | $0.045 | $0.14 | $0.045 / $0.14 | $0.069 | 1.31M | on InferenceNet (fp4) | · |
Z.ai · 1.31M | $0.561 | $1.76 | $0.561 / $1.76 | $0.862 | 1.31M | on Baidu (fp8) | · |
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75100% | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | 100% |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | · |
Moonshot AI · 262K | $0.45 | $2.25 | $0.45 / $2.25 | $0.9 | 262K | on SiliconFlow (int4) | · |
About GLM 4.6
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
Takes text in and returns text. First listed Sep 23, 2026. Up to 16K output tokens. Blended price $0.76 per 1M tokens at 3 input to 1 output. Overall score 5 on leadersboard.fru.dev (rank 30). Model id glm-4-6.
Price history
No list-price history: this model is sold by hosts rather than its lab.
Across fru.devCompany profile
Questions
How much does GLM 4.6 cost?
$0.43 per million input tokens and $1.75 per million output tokens, with cached input at $0.08.
What is the context window of GLM 4.6?
205K tokens, with up to 16K tokens of output.
Where is GLM 4.6 cheapest?
Of 4 routes tracked, Venice is cheapest at $0.43 input and $1.75 output per million tokens.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.