Z.ai
GLM 5 API prices by provider
z-ai/glm-5
8 inference providers serve this exact model through OpenRouter. Input and output rates are paired per endpoint; route tags may identify a region or sub-endpoint whose eligibility is not yet verified.
- Active endpoints
- 8 / 8
- Lowest at 1M in + 1M out
- $2.94
- Baidu · standard tier
- Context up to
- 205K
- Latest price change
- None yet
- Rates unchanged since first record 2026-09-29
| Provider · route | Tier | Input / 1M | Output / 1M | Cache read | Cache write | Context | Tools | 1M + 1M | Status | Price since |
|---|---|---|---|---|---|---|---|---|---|---|
GMICloud gmicloud/fp8 · fp8 | Promo | $0.60 | $1.92 | $0.12 | — | 203K | Yes | $2.52 | Active | Since |
StreamLake streamlake/fp8 · fp8 | Promo | $0.60 | $1.92 | $0.12 | — | 198K | Yes | $2.52 | Active | Since |
Baidu baidu/fp8 · fp8 | Standard | $0.70 | $2.24 | $0.14 | — | 203K | Yes | $2.94 | Active | Since |
SiliconFlow siliconflow/fp8 · fp8 | Standard | $0.95 | $2.55 | $0.20 | — | 205K | Yes | $3.50 | Active | Since |
Amazon Bedrock amazon-bedrock · unknown | Standard | $1.00 | $3.20 | — | — | 203K | Yes | $4.20 | Active | Since |
Novita novita/fp8 · fp8 | Standard | $1.00 | $3.20 | $0.20 | — | 203K | Yes | $4.20 | Active | Since |
Venice venice/fp8 · fp8 | Standard | $1.00 | $3.20 | $0.20 | — | 198K | Yes | $4.20 | Active | Since |
Z.AI z-ai/fp8 · fp8 | Standard | $1.00 | $3.20 | $0.20 | — | 203K | Yes | $4.20 | Active | Since |
“—” means no rate is published for that endpoint; it is excluded from estimates, not counted as zero. “Since” is the first record with no rate change observed; “Changed” is the latest observed rate or billing change. Per-request fees: not published for any endpoint here. Status is the upstream operational flag, not a measured uptime.