Skip to content

GLM 4.7 Flash

Z.ai · Open weights · 200K context · released Jan 19, 2026 · checked 7h ago

Ways to buy 3

Way to buyIn / out per 1M
Venice
fp8 · cached $0.01
$0.06 / $0.4Cheapest
Cloudflare Workers AI
Default region
$0.06 / $0.4Cheapest
Novita AI
bf16 · cached $0.01
$0.07 / $0.4+5% vs cheapest

USD per 1M tokens. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $14.00 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
GLM 5.3 Prime
Z.ai · 1M
$2.8 / $8.8
GLM 5.3 FlashX
Z.ai · 1M
$0.37 / $1.25
GLM 5.3 Flash
Z.ai · 1.31M
$0.045 / $0.14
GLM 5.3
Z.ai · 1.31M
$0.561 / $1.76
DeepSeek V4.1 Flash
DeepSeek · 1MCheapest: $0.131 on Morph
$0.15 / $0.6
GPT-6 Luna
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex)
$0.1 / $0.5
GPT-6 Luna Pro
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex)
$0.1 / $0.5
Qwen3.8 Omni Flash
Qwen (Alibaba) · 1M
$0.15 / $0.47

About GLM 4.7 Flash

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency.

Takes text in and returns text. First listed Sep 23, 2026. Up to 118K output tokens. Blended price $0.145 per 1M tokens at 3 input to 1 output. Model id glm-4-7-flash.

Price history

No list-price history: this model is sold by hosts rather than its lab.

Across fru.devCompany profile

Questions

How much does GLM 4.7 Flash cost?

$0.06 per million input tokens and $0.4 per million output tokens, with cached input at $0.01.

What is the context window of GLM 4.7 Flash?

200K tokens, with up to 118K tokens of output.

Where is GLM 4.7 Flash cheapest?

Of 3 routes tracked, Venice is cheapest at $0.06 input and $0.4 output per million tokens.

Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.