FindCostShortlistChangesModels
← All tracked models

Z.ai

GLM 5 API prices by provider

z-ai/glm-5

8 inference providers serve this exact model through OpenRouter. Input and output rates are paired per endpoint; route tags may identify a region or sub-endpoint whose eligibility is not yet verified.

Active endpoints
8 / 8
Lowest at 1M in + 1M out
$2.94
Baidu · standard tier
Context up to
205K
Latest price change
None yet
Rates unchanged since first record 2026-09-29
GLM 5 endpoint prices per million tokens
Provider · routeTierInput / 1MOutput / 1MCache readCache writeContextTools1M + 1MStatusPrice since

GMICloud

gmicloud/fp8 · fp8

Promo$0.60$1.92$0.12—203KYes$2.52ActiveSince

StreamLake

streamlake/fp8 · fp8

Promo$0.60$1.92$0.12—198KYes$2.52ActiveSince

Baidu

baidu/fp8 · fp8

Standard$0.70$2.24$0.14—203KYes$2.94ActiveSince

SiliconFlow

siliconflow/fp8 · fp8

Standard$0.95$2.55$0.20—205KYes$3.50ActiveSince

Amazon Bedrock

amazon-bedrock · unknown

Standard$1.00$3.20——203KYes$4.20ActiveSince

Novita

novita/fp8 · fp8

Standard$1.00$3.20$0.20—203KYes$4.20ActiveSince

Venice

venice/fp8 · fp8

Standard$1.00$3.20$0.20—198KYes$4.20ActiveSince

Z.AI

z-ai/fp8 · fp8

Standard$1.00$3.20$0.20—203KYes$4.20ActiveSince

“—” means no rate is published for that endpoint; it is excluded from estimates, not counted as zero. “Since” is the first record with no rate change observed; “Changed” is the latest observed rate or billing change. Per-request fees: not published for any endpoint here. Status is the upstream operational flag, not a measured uptime.