GLM 4.5V
z-ai/glm-4.5v
Compare with GLM 4.5 Air, GLM 4.5, GLM 4.6
CompareGLM 4.5V
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
Novitanovita/fp8fp8 | No data | — | 98.21% | 1.59s | 61.6s | 39 tps | 66K | $0.600 / $1.80 |
Z.AIz-ai/fp8fp8 | Degraded | — | 96.34% | 2.58s | 21.3s | 39 tps | 66K | $0.600 / $1.80 |
- Novita—No data24h 98.21%
- p50
- 1.59s
- p99
- 61.6s
- Throughput
- 39 tps
- Context
- 66K
- $/Mtok
- $0.600 / $1.80
novita/fp8fp8 - Z.AI—Degraded24h 96.34%
- p50
- 2.58s
- p99
- 21.3s
- Throughput
- 39 tps
- Context
- 66K
- $/Mtok
- $0.600 / $1.80
z-ai/fp8fp8
Measured 3m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
1 endpoint reporting no data is omitted from the charts.
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows, averaged over the hour.
Throughput
Median tokens per second.