# Z.ai: GLM 4.7 — price and measured latency by host

Updated July 25, 2026. Served by 9 providers; 5 of them measured here.

Latency and uptime are measured by llmlatency.dev probes. Prices are published by the providers and read via OpenRouter on 2026-07-25 — not measured.

| Host | Input $/M | Output $/M | Context | Fastest p50 | From | Uptime |
| --- | --- | --- | --- | --- | --- | --- |
| deepinfra | $0.400 | $1.750 | 202,752 | 226 ms | us-central | 99.9% |
| novita | $0.540 | $1.980 | 204,800 | 71 ms | us-central | 100% |
| glm | $0.600 | $2.200 | 202,752 | 349 ms | ap-tokyo | 100% |
| google | $0.600 | $2.200 | 200,000 | 40 ms | us-central | 100% |
| cerebras | $2.250 | $2.750 | 131,072 | 69 ms | us-central | 100% |

Source: https://llmlatency.dev/model/z-ai-glm-4.7 · CC BY 4.0
