GLM 4.5V
z-ai/glm-4.5v
Compare with GLM 4.5 Air, GLM 4.5, GLM 4.6
CompareGLM 4.5V
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
Novitanovita/fp8fp8 | No data | — | 98.46% | 3.88s | 47.8s | 55 tps | 66K | $0.600 / $1.80 |
Z.AIz-ai/fp8fp8 | Degraded | — | 96.70% | 2.69s | 36.8s | 53 tps | 66K | $0.600 / $1.80 |
- Novita—No data24h 98.46%
- p50
- 3.88s
- p99
- 47.8s
- Throughput
- 55 tps
- Context
- 66K
- $/Mtok
- $0.600 / $1.80
novita/fp8fp8 - Z.AI—Degraded24h 96.70%
- p50
- 2.69s
- p99
- 36.8s
- Throughput
- 53 tps
- Context
- 66K
- $/Mtok
- $0.600 / $1.80
z-ai/fp8fp8
Measured 5m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
1 endpoint reporting no data is omitted from the charts.
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows.
Throughput
Median tokens per second.