FindCostShortlistChangesModels
← All tracked models

Z.ai

GLM 4.7 Flash API prices by provider

z-ai/glm-4.7-flash

3 inference providers serve this exact model through OpenRouter. Input and output rates are paired per endpoint; route tags may identify a region or sub-endpoint whose eligibility is not yet verified.

Active endpoints
2 / 3
Lowest at 1M in + 1M out
$0.46
Venice · standard tier
Context up to
200K
Latest price change
None yet
Rates unchanged since first record 2026-09-29
GLM 4.7 Flash endpoint prices per million tokens
Provider · routeTierInput / 1MOutput / 1MCache readCache writeContextTools1M + 1MStatusPrice since

Venice

venice/fp8 · fp8

Standard$0.06$0.40$0.01—128KYes$0.46ActiveSince

Cloudflare

cloudflare · unknown

Standard$0.0605$0.40——131KYes$0.4605ActiveSince

Novita

novita/bf16 · bf16

Standard$0.07$0.40$0.01—200KYes$0.47UnavailableSince

“—” means no rate is published for that endpoint; it is excluded from estimates, not counted as zero. “Since” is the first record with no rate change observed; “Changed” is the latest observed rate or billing change. Per-request fees: not published for any endpoint here. Status is the upstream operational flag, not a measured uptime.