modelstatus.dev

GLM 5

z-ai/glm-5InputtextOutputtext

Compare its 11 providersIncident RSSCompare with GLM 5 Turbo, GLM 5V Turbo, GLM 5.1

CompareGLM 5
  • Healthy24h 99.04%
    7d
    98.47%
    30d
    98.50%
    p50
    2.78s
    p99
    7.11s
    Throughput
    54 tps
    Context
    203K
    $/Mtok
    $0.600 / $1.92
    gmicloud/fp8fp8
  • Novita100%
    Healthy24h 99.99%
    7d
    100%
    30d
    99.97%
    p50
    1.31s
    p99
    11.3s
    Throughput
    45 tps
    Context
    203K
    $/Mtok
    $1.00 / $3.20
    novita/fp8fp8
  • Z.AI100%
    Healthy24h 99.99%
    7d
    99.95%
    30d
    99.93%
    p50
    7.24s
    p99
    13.4s
    Throughput
    50 tps
    Context
    203K
    $/Mtok
    $1.00 / $3.20
    z-ai/fp8fp8
  • Healthy24h 99.72%
    7d
    99.77%
    30d
    99.73%
    p50
    2.59s
    p99
    8.14s
    Throughput
    46 tps
    Context
    198K
    $/Mtok
    $0.600 / $1.92
    streamlake/fp8fp8
  • No data24h 98.58%
    7d
    99.43%
    30d
    97.75%
    p50
    786ms
    p99
    6.05s
    Throughput
    87 tps
    Context
    203K
    $/Mtok
    $1.00 / $3.20
    amazon-bedrock
  • No data24h —
    7d
    —
    30d
    —
    p50
    —
    p99
    —
    Throughput
    —
    Context
    203K
    $/Mtok
    $0.950 / $3.15
    atlas-cloud/fp8fp8
  • Degraded24h 99.47%
    7d
    99.16%
    30d
    99.63%
    p50
    1.44s
    p99
    26.1s
    Throughput
    44 tps
    Context
    203K
    $/Mtok
    $0.700 / $2.24
    baidu/fp8fp8
  • No data24h —
    7d
    —
    30d
    99.88%
    p50
    —
    p99
    —
    Throughput
    —
    Context
    203K
    $/Mtok
    $0.600 / $2.08
    deepinfra/fp4fp4
  • No data24h —
    7d
    —
    30d
    95.05%
    p50
    —
    p99
    —
    Throughput
    —
    Context
    64K
    $/Mtok
    $1.00 / $3.20
    digitalocean
  • No data24h 99.69%
    7d
    99.68%
    30d
    99.45%
    p50
    2.53s
    p99
    6.29s
    Throughput
    40 tps
    Context
    205K
    $/Mtok
    $0.950 / $2.55
    siliconflow/fp8fp8
  • No data24h 99.71%
    7d
    99.11%
    30d
    98.53%
    p50
    1.47s
    p99
    18.0s
    Throughput
    74 tps
    Context
    198K
    $/Mtok
    $1.00 / $3.20
    venice/fp8fp8

Measured 3m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.

3 endpoints have no chartable data in this range.

Availability

Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.

Time to first token

Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.

Throughput

Median tokens per second.

Daily availability

One row per endpoint · one dot per day · last 90 days.

worst hour ≥ 95%90–95%< 90%no data

Colour by each endpoint's worst hour of each UTC day · 11 endpoints · 47 days with data, 76 endpoint-days below 90%.

Price changes

  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-09-26 00:50 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-09-26 00:40 UTC
  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-09-09 10:10 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-09-09 10:05 UTC
  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-08-27 18:45 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-08-27 18:20 UTC
  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-08-26 18:31 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-08-26 14:36 UTC
  • GLM 5GMICloudgmicloud/fp8$1.00 → $0.600 / $3.20 → $1.92cache read $0.200 → $0.1202026-08-26 13:51 UTC
  • GLM 5GMICloudgmicloud/fp8$0.600 → $1.00 / $1.92 → $3.20cache read $0.120 → $0.2002026-08-26 13:46 UTC
  • GLM 5DigitalOceandigitalocean$0.750 → $1.00 / $2.40 → $3.202026-08-21 03:26 UTC