FindCostShortlistChangesModels
← All tracked models

Z.ai

GLM 4.6 API prices by provider

z-ai/glm-4.6

4 inference providers serve this exact model through OpenRouter. Input and output rates are paired per endpoint; route tags may identify a region or sub-endpoint whose eligibility is not yet verified.

Active endpoints
4 / 4
Lowest at 1M in + 1M out
$2.18
Venice · standard tier
Context up to
205K
Latest price change
None yet
Rates unchanged since first record 2026-09-29
GLM 4.6 endpoint prices per million tokens
Provider · routeTierInput / 1MOutput / 1MCache readCache writeContextTools1M + 1MStatusPrice since

Venice

venice/fp4 · fp4

Standard$0.43$1.75$0.08—198KYes$2.18ActiveSince

DeepInfra

deepinfra/fp4 · fp4

Standard$0.50$2.00$0.10—203KYes$2.50ActiveSince

Novita

novita/bf16 · bf16

Standard$0.55$2.20$0.11—205KYes$2.75ActiveSince

Z.AI

z-ai/fp4 · fp4

Standard$0.60$2.20$0.11—203KYes$2.80ActiveSince

“—” means no rate is published for that endpoint; it is excluded from estimates, not counted as zero. “Since” is the first record with no rate change observed; “Changed” is the latest observed rate or billing change. Per-request fees: not published for any endpoint here. Status is the upstream operational flag, not a measured uptime.