FindCostShortlistChangesModels
← All tracked models

Z.ai

GLM 4.7 API prices by provider

z-ai/glm-4.7

7 inference providers serve this exact model through OpenRouter. Input and output rates are paired per endpoint; route tags may identify a region or sub-endpoint whose eligibility is not yet verified.

Active endpoints
3 / 7
Lowest at 1M in + 1M out
$2.80
Google · standard tier
Context up to
205K
Latest price change
None yet
Rates unchanged since first record 2026-09-29
GLM 4.7 endpoint prices per million tokens
Provider · routeTierInput / 1MOutput / 1MCache readCache writeContextTools1M + 1MStatusPrice since

Novita

novita/fp8 · fp8

Promo$0.54$1.98$0.099—205KYes$2.52ActiveSince

Google

google-vertex · unknown

Standard$0.60$2.20——200KYes$2.80ActiveSince

Z.AI

z-ai/fp4 · fp4

Standard$0.60$2.20$0.11—203KYes$2.80ActiveSince

DeepInfra

deepinfra/fp4 · fp4

Standard$0.40$1.75$0.08—203KYes$2.15UnavailableSince

Venice

venice/fp4 · fp4

Promo$0.4004$1.9292$0.0801—198KYes$2.3296UnavailableSince

AtlasCloud

atlas-cloud/fp8 · fp8

Standard$0.52$1.85$0.12—203KYes$2.37UnavailableSince

Mancer 2

mancer/fp4 · fp4

Standard$0.70$2.50——131KNot declared$3.20UnavailableSince

“—” means no rate is published for that endpoint; it is excluded from estimates, not counted as zero. “Since” is the first record with no rate change observed; “Changed” is the latest observed rate or billing change. Per-request fees: not published for any endpoint here. Status is the upstream operational flag, not a measured uptime.