Skip to content

GLM 5.2

Z.ai · Open weights · 1M context · released Jun 16, 2026 · checked 6h ago

Ways to buy 27

Way to buyIn / out per 1M
Baidu
fp8 · cached $0.104
$0.561 / $1.76Cheapest
DeepInfra
fp4 · cached $0.105
$0.563 / $1.8+1% vs cheapest
StreamLake
fp8 · cached $0.119
$0.641 / $2.01+14% vs cheapest
Novita AI
fp8 · cached $0.121
$0.65 / $2.04+16% vs cheapest
Baidu
fast, fp4 · cached $0.161
$0.648 / $2.27+22% vs cheapest
DigitalOcean
Default region · cached $0.105
$0.7 / $2.2+25% vs cheapest
CoreWeave
fp4 · cached $0.14
$0.76 / $2.42+36% vs cheapest
AtlasCloud
fp8 · cached $0.174
$0.938 / $2.95+67% vs cheapest
Inceptron
fp4 · cached $0.188
$0.898 / $3.18+70% vs cheapest
Qwen (Alibaba)
fp8 · cached $0.193
$0.966 / $3.04+72% vs cheapest
Phala
fp8 · cached $0.22
$1.26 / $3+97% vs cheapest
SiliconFlow
fp8 · cached $0.221
$1.19 / $3.74+112% vs cheapest
Baseten
fp8 · cached $0.14
$1.4 / $4.4+149% vs cheapest
Cloudflare Workers AI
Default region · cached $0.26
$1.4 / $4.4+149% vs cheapest
Fireworks AI
Default region · cached $0.14
$1.4 / $4.4+149% vs cheapest
Friendli
Default region · cached $0.26
$1.4 / $4.4+149% vs cheapest
GMICloud
fp8 · cached $0.26
$1.4 / $4.4+149% vs cheapest
Parasail
fp4 · cached $0.26
$1.4 / $4.4+149% vs cheapest
Together AI
Default region · cached $0.26
$1.4 / $4.4+149% vs cheapest
Venice
fp8 · cached $0.26
$1.4 / $4.4+149% vs cheapest
Z.ai
fp8 · cached $0.26
$1.4 / $4.4+149% vs cheapest
Wafer
Default region · cached $0.26
$1.19 / $5.3+157% vs cheapest
Baseten
fast, fp8 · cached $0.21
$2.1 / $6.6+274% vs cheapest
Fireworks AI
fast · cached $0.21
$2.1 / $6.6+274% vs cheapest
Fireworks AI
fast-us · cached $0.21
$2.1 / $6.6+274% vs cheapest
Qwen (Alibaba)
fast, fp8 · cached $0.462
$2.31 / $7.26+311% vs cheapest
Decart
fast, fp4 · cached $0.48
$2.25 / $8+328% vs cheapest

USD per 1M tokens. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $91.43 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
GLM 5.3 Prime
Z.ai · 1M
$2.8 / $8.8
GLM 5.3 FlashX
Z.ai · 1M
$0.37 / $1.25
GLM 5.3 Flash
Z.ai · 1.31M
$0.045 / $0.14
GLM 5.3
Z.ai · 1.31M
$0.561 / $1.76
Gemini 3.8 Flash
Google · 1MCheapest: $0.75 on Google Gemini API (flex)
$0.75 / $3.75
Gemini 3.7 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75100%
Gemini 3.6 Flash
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex)
$0.75 / $3.75
Kimi K2.5
Moonshot AI · 262K
$0.45 / $2.25

About GLM 5.2

GLM 5.2 is a large-scale reasoning model from Z.ai.

Takes text in and returns text. First listed Sep 23, 2026. Up to 131K output tokens. Blended price $0.862 per 1M tokens at 3 input to 1 output. Model id glm-5-2.

Price history

No list-price history: this model is sold by hosts rather than its lab.

Across fru.devCompany profile

Questions

How much does GLM 5.2 cost?

$0.561 per million input tokens and $1.76 per million output tokens, with cached input at $0.104.

What is the context window of GLM 5.2?

1M tokens, with up to 131K tokens of output.

Where is GLM 5.2 cheapest?

Of 27 routes tracked, Baidu is cheapest at $0.561 input and $1.76 output per million tokens.

Has GLM 5.2's price changed?

Yes. The output price went from $3.24 to $3.18, first seen Sep 25, 2026.

Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.