modelstatus.dev

GLM 4.6

z-ai/glm-4.6

Compare with GLM 4.6V, GLM 4.5 Air, GLM 4.5V

Compare its 5 providers · Incident RSS

Showing only the venice/fp4 provider · show all 5

CompareGLM 4.6
  • Venice100%
    Healthy24h 99.35%
    7d
    98.86%
    30d
    98.86%
    p50
    1.16s
    p99
    17.0s
    Throughput
    18 tps
    Context
    198K
    $/Mtok
    $0.430 / $1.75
    venice/fp4fp4

Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.

Availability

Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.

Time to first token

Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window, averaged over the hour.

Throughput

Median tokens per second.

Daily availability

One row per endpoint · one cell per week · up to a year.

worst hour ≥ 95%90–95%< 90%no data

Colour by each endpoint's worst hour of each UTC week · 5 endpoints · 24 weeks with data, 13 endpoint-weeks below 90%.