GLM 5.3 (batch)
z-ai/glm-5.3:batchInputOutput
Compare its 2 providersIncident RSSCompare with GLM 5.3, GLM 5.3 Flash, GLM 5.3 FlashX
Showing only the deepinfra/fp4 provider · show all 2
CompareGLM 5.3 (batch)
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
| No data | — | — | — | — | — | — | — | 1.0M | $0.450 / $2.00 |
- No data24h —
- 7d
- —
- 30d
- —
- p50
- —
- p99
- —
- Throughput
- —
- Context
- 1.0M
- $/Mtok
- $0.450 / $2.00
deepinfra/fp4fp4
Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
1 endpoint has no chartable data in this range.
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.
Throughput
Median tokens per second.
Daily availability
One row per endpoint · one dot per day · last 90 days.
worst hour ≥ 95%90–95%< 90%no data
Colour by each endpoint's worst hour of each UTC day · 2 endpoints · 0 days with data, none below the red threshold. Tables and charts above are filtered to deepinfra/fp4.
Price changes
- GLM 5.3 (batch)DeepInfradeepinfra/fp4$0.600 → $0.450 / $2.002026-09-23 15:16 UTC
- GLM 5.3 (batch)DeepInfradeepinfra/fp4$0.720 → $0.600 / $2.40 → $2.00cache read $0.120 → $0.1002026-09-23 15:11 UTC