Ways to buy 27
| Way to buy | Input | Output | In / out per 1M | Cached | Vs cheapest |
|---|---|---|---|---|---|
| Baidu fp8 · cached $0.104 | $0.561 | $1.76 | $0.561 / $1.76Cheapest | $0.104 | Cheapest |
fp4 · cached $0.105 | $0.563 | $1.8 | $0.563 / $1.8+1% vs cheapest | $0.105 | +1% vs cheapest |
| StreamLake fp8 · cached $0.119 | $0.641 | $2.01 | $0.641 / $2.01+14% vs cheapest | $0.119 | +14% vs cheapest |
fp8 · cached $0.121 | $0.65 | $2.04 | $0.65 / $2.04+16% vs cheapest | $0.121 | +16% vs cheapest |
| Baidu fast, fp4 · cached $0.161 | $0.648 | $2.27 | $0.648 / $2.27+22% vs cheapest | $0.161 | +22% vs cheapest |
Default region · cached $0.105 | $0.7 | $2.2 | $0.7 / $2.2+25% vs cheapest | $0.105 | +25% vs cheapest |
fp4 · cached $0.14 | $0.76 | $2.42 | $0.76 / $2.42+36% vs cheapest | $0.14 | +36% vs cheapest |
| AtlasCloud fp8 · cached $0.174 | $0.938 | $2.95 | $0.938 / $2.95+67% vs cheapest | $0.174 | +67% vs cheapest |
| Inceptron fp4 · cached $0.188 | $0.898 | $3.18 | $0.898 / $3.18+70% vs cheapest | $0.188 | +70% vs cheapest |
fp8 · cached $0.193 | $0.966 | $3.04 | $0.966 / $3.04+72% vs cheapest | $0.193 | +72% vs cheapest |
| Phala fp8 · cached $0.22 | $1.26 | $3 | $1.26 / $3+97% vs cheapest | $0.22 | +97% vs cheapest |
fp8 · cached $0.221 | $1.19 | $3.74 | $1.19 / $3.74+112% vs cheapest | $0.221 | +112% vs cheapest |
fp8 · cached $0.14 | $1.4 | $4.4 | $1.4 / $4.4+149% vs cheapest | $0.14 | +149% vs cheapest |
Default region · cached $0.26 | $1.4 | $4.4 | $1.4 / $4.4+149% vs cheapest | $0.26 | +149% vs cheapest |
Default region · cached $0.14 | $1.4 | $4.4 | $1.4 / $4.4+149% vs cheapest | $0.14 | +149% vs cheapest |
Default region · cached $0.26 | $1.4 | $4.4 | $1.4 / $4.4+149% vs cheapest | $0.26 | +149% vs cheapest |
| GMICloud fp8 · cached $0.26 | $1.4 | $4.4 | $1.4 / $4.4+149% vs cheapest | $0.26 | +149% vs cheapest |
fp4 · cached $0.26 | $1.4 | $4.4 | $1.4 / $4.4+149% vs cheapest | $0.26 | +149% vs cheapest |
Default region · cached $0.26 | $1.4 | $4.4 | $1.4 / $4.4+149% vs cheapest | $0.26 | +149% vs cheapest |
fp8 · cached $0.26 | $1.4 | $4.4 | $1.4 / $4.4+149% vs cheapest | $0.26 | +149% vs cheapest |
fp8 · cached $0.26 | $1.4 | $4.4 | $1.4 / $4.4+149% vs cheapest | $0.26 | +149% vs cheapest |
| Wafer Default region · cached $0.26 | $1.19 | $5.3 | $1.19 / $5.3+157% vs cheapest | $0.26 | +157% vs cheapest |
fast, fp8 · cached $0.21 | $2.1 | $6.6 | $2.1 / $6.6+274% vs cheapest | $0.21 | +274% vs cheapest |
fast · cached $0.21 | $2.1 | $6.6 | $2.1 / $6.6+274% vs cheapest | $0.21 | +274% vs cheapest |
fast-us · cached $0.21 | $2.1 | $6.6 | $2.1 / $6.6+274% vs cheapest | $0.21 | +274% vs cheapest |
fast, fp8 · cached $0.462 | $2.31 | $7.26 | $2.31 / $7.26+311% vs cheapest | $0.462 | +311% vs cheapest |
| Decart fast, fp4 · cached $0.48 | $2.25 | $8 | $2.25 / $8+328% vs cheapest | $0.48 | +328% vs cheapest |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $91.43 direct.
Helicone$91.43+0%
Kong AI Gateway$91.43+0%
LiteLLM$91.43+0%
Vercel AI Gateway$91.43+0%
Cloudflare AI Gateway$96.00+5%
Requesty$96.00+5%
Eden AI$96.46+5.5%
OpenRouter$96.46+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Z.ai · 1M | $2.8 | $8.8 | $2.8 / $8.8 | $4.3 | 1M | Lab list price | · |
Z.ai · 1M | $0.37 | $1.25 | $0.37 / $1.25 | $0.59 | 1M | Lab list price | · |
Z.ai · 1.31M | $0.045 | $0.14 | $0.045 / $0.14 | $0.069 | 1.31M | on InferenceNet (fp4) | · |
Z.ai · 1.31M | $0.561 | $1.76 | $0.561 / $1.76 | $0.862 | 1.31M | on Baidu (fp8) | · |
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75100% | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | 100% |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | · |
Moonshot AI · 262K | $0.45 | $2.25 | $0.45 / $2.25 | $0.9 | 262K | on SiliconFlow (int4) | · |
About GLM 5.2
GLM 5.2 is a large-scale reasoning model from Z.ai.
Takes text in and returns text. First listed Sep 23, 2026. Up to 131K output tokens. Blended price $0.862 per 1M tokens at 3 input to 1 output. Model id glm-5-2.
Price history
No list-price history: this model is sold by hosts rather than its lab.
Across fru.devCompany profile
Questions
How much does GLM 5.2 cost?
$0.561 per million input tokens and $1.76 per million output tokens, with cached input at $0.104.
What is the context window of GLM 5.2?
1M tokens, with up to 131K tokens of output.
Where is GLM 5.2 cheapest?
Of 27 routes tracked, Baidu is cheapest at $0.561 input and $1.76 output per million tokens.
Has GLM 5.2's price changed?
Yes. The output price went from $3.24 to $3.18, first seen Sep 25, 2026.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.