Skip to content

Compare model prices

Input, output, cached and batch prices, every route and a month of tokens, side by side.

GLM 4.6GLM 5.3 Prime

A month of tokens

  • GLM 4.6$78.00
  • GLM 5.3 Prime$456.00

100M input and 20M output tokens, list price. Use your own numbers.

Side by side

GLM 4.6GLM 5.3 Prime
Input per 1M$0.43$2.8
Output per 1M$1.75$8.8
Cached input$0.08$0.56
Batch in / outn/an/a
Blended (3:1)$0.76$4.3
Context205K1M
Routes42
ReleasedSep 30, 2025Sep 23, 2026
Leadersboard score5 (#30)n/a

On the price map

Size: context window, 8K to 2MEqual blended cost (3:1) 245 models, lower left is cheaper

Every route

GLM 4.6

Way to buyIn / out per 1M
Venice
fp4 · cached $0.08
$0.43 / $1.75Same as list
DeepInfra
fp4 · cached $0.1
$0.5 / $2+15% vs list
Novita AI
bf16 · cached $0.11
$0.55 / $2.2+27% vs list
Z.ai
fp4 · cached $0.11
$0.6 / $2.2+32% vs list

GLM 5.3 Prime

Way to buyIn / out per 1M
Z.ai API
The lab's own API · cached $0.56
$2.8 / $8.8List price
Qwen (Alibaba)
Default region · cached $0.56
$2.8 / $8.8Same as list

Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.