modelstatus.dev

GLM 5

z-ai/glm-5InputtextOutputtext

Compare its 11 providersIncident RSSCompare with GLM 5 Turbo, GLM 5V Turbo, GLM 5.1

CompareGLM 5
  • Healthy24h 98.49%
    7d
    99.42%
    30d
    97.75%
    p50
    964ms
    p99
    10.5s
    Throughput
    80 tps
    Context
    203K
    $/Mtok
    $1.00 / $3.20
    amazon-bedrock
  • Healthy24h 99.15%
    7d
    98.40%
    30d
    98.50%
    p50
    2.47s
    p99
    7.46s
    Throughput
    57 tps
    Context
    203K
    $/Mtok
    $0.600 / $1.92
    gmicloud/fp8fp8
  • Novita100%
    Healthy24h 99.99%
    7d
    100%
    30d
    99.97%
    p50
    1.21s
    p99
    6.49s
    Throughput
    43 tps
    Context
    203K
    $/Mtok
    $1.00 / $3.20
    novita/fp8fp8
  • Healthy24h 99.70%
    7d
    99.69%
    30d
    99.46%
    p50
    2.45s
    p99
    10.0s
    Throughput
    45 tps
    Context
    205K
    $/Mtok
    $0.950 / $2.55
    siliconflow/fp8fp8
  • Healthy24h 99.70%
    7d
    99.77%
    30d
    99.74%
    p50
    2.21s
    p99
    7.10s
    Throughput
    52 tps
    Context
    198K
    $/Mtok
    $0.600 / $1.92
    streamlake/fp8fp8
  • Z.AI100%
    Healthy24h 99.93%
    7d
    99.95%
    30d
    99.93%
    p50
    7.77s
    p99
    12.7s
    Throughput
    49 tps
    Context
    203K
    $/Mtok
    $1.00 / $3.20
    z-ai/fp8fp8
  • No data24h —
    7d
    —
    30d
    —
    p50
    —
    p99
    —
    Throughput
    —
    Context
    203K
    $/Mtok
    $0.950 / $3.15
    atlas-cloud/fp8fp8
  • No data24h 99.16%
    7d
    99.15%
    30d
    99.63%
    p50
    1.08s
    p99
    6.92s
    Throughput
    56 tps
    Context
    203K
    $/Mtok
    $0.700 / $2.24
    baidu/fp8fp8
  • No data24h —
    7d
    —
    30d
    99.88%
    p50
    —
    p99
    —
    Throughput
    —
    Context
    203K
    $/Mtok
    $0.600 / $2.08
    deepinfra/fp4fp4
  • No data24h —
    7d
    —
    30d
    95.21%
    p50
    —
    p99
    —
    Throughput
    —
    Context
    64K
    $/Mtok
    $1.00 / $3.20
    digitalocean
  • No data24h 99.65%
    7d
    99.12%
    30d
    98.52%
    p50
    1.33s
    p99
    8.51s
    Throughput
    52 tps
    Context
    198K
    $/Mtok
    $1.00 / $3.20
    venice/fp8fp8

Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.

3 endpoints have no chartable data in this range.

Availability

Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.

Time to first token

Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.

Throughput

Median tokens per second.

Daily availability

One row per endpoint · one dot per day · last 90 days.

worst hour ≥ 95%90–95%< 90%no data

Colour by each endpoint's worst hour of each UTC day · 11 endpoints · 46 days with data, 79 endpoint-days below 90%.

Price changes

  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-09-26 00:50 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-09-26 00:40 UTC
  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-09-09 10:10 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-09-09 10:05 UTC
  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-08-27 18:45 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-08-27 18:20 UTC
  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-08-26 18:31 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-08-26 14:36 UTC
  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-08-26 13:51 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-08-26 13:46 UTC
  • GLM 5DigitalOceandigitalocean$0.750 → $1.00 / $2.40 → $3.202026-08-21 03:26 UTC