modelstatus.dev

GLM 5.1

GLM 5.1: providers compared

Every endpoint serving this model, side by side. ◆ marks the best value in each column — per model, never fleet-wide.

  • AtlasCloud100% (best)
    Healthy24h 99.69%
    7d
    99.75%
    30d
    99.75%
    p50
    1.53s
    p99
    13.8s
    Throughput
    55 tps
    Context
    203K
    $/Mtok
    $1.26 / $3.96
    atlas-cloud/fp8fp8
  • Baidu100% (best)
    Healthy24h 99.97%
    7d
    99.90%
    30d
    99.90%
    p50
    1.28s
    p99
    4.72s
    Throughput
    60 tps
    Context
    203K
    $/Mtok
    $0.896 / $2.82 (best)
    baidu/fp8fp8
  • DigitalOcean100% (best)
    Healthy24h 97.31%
    7d
    94.07%
    30d
    94.07%
    p50
    1.39s
    p99
    22.1s
    Throughput
    14 tps
    Context
    164K
    $/Mtok
    $1.30 / $4.30
    digitalocean
  • Friendli100% (best)
    Healthy24h 99.98%
    7d
    99.96% (best)
    30d
    99.96% (best)
    p50
    239ms (best)
    p99
    3.69s (best)
    Throughput
    74 tps
    Context
    203K
    $/Mtok
    $1.40 / $4.40
    friendli
  • Nebius100% (best)
    Healthy24h 98.21%
    7d
    97.51%
    30d
    97.51%
    p50
    812ms
    p99
    8.70s
    Throughput
    28 tps
    Context
    203K
    $/Mtok
    $1.40 / $4.40
    nebius/fp8fp8
  • Parasail100% (best)
    Healthy24h 98.91%
    7d
    99.43%
    30d
    99.43%
    p50
    839ms
    p99
    13.1s
    Throughput
    89 tps (best)
    Context
    203K
    $/Mtok
    $1.40 / $4.40
    parasail/fp8fp8
  • SiliconFlow100% (best)
    Healthy24h 99.95%
    7d
    99.89%
    30d
    99.89%
    p50
    1.74s
    p99
    10.2s
    Throughput
    48 tps
    Context
    205K
    $/Mtok
    $1.19 / $3.74
    siliconflow/fp8fp8
  • StreamLake100% (best)
    Healthy24h 99.51%
    7d
    98.83%
    30d
    98.83%
    p50
    2.31s
    p99
    9.94s
    Throughput
    28 tps
    Context
    200K
    $/Mtok
    $0.966 / $3.04
    streamlake/fp8fp8
  • Z.AI100% (best)
    Healthy24h 99.89%
    7d
    99.09%
    30d
    99.09%
    p50
    3.96s
    p99
    6.71s
    Throughput
    23 tps
    Context
    203K
    $/Mtok
    $1.40 / $4.40
    z-ai/fp8fp8
  • GMICloud97.73%
    Healthy24h 99.64%
    7d
    99.63%
    30d
    99.63%
    p50
    1.71s
    p99
    13.7s
    Throughput
    63 tps
    Context
    203K
    $/Mtok
    $0.910 / $2.86
    gmicloud/fp8fp8
  • No data24h 99.90%
    7d
    99.77%
    30d
    99.77%
    p50
    1.91s
    p99
    9.35s
    Throughput
    48 tps
    Context
    203K
    $/Mtok
    $1.33 / $4.18
    alibaba/fp8fp8
  • Down24h 76.88%
    7d
    83.27%
    30d
    83.27%
    p50
    3.26s
    p99
    53.1s
    Throughput
    25 tps
    Context
    203K
    $/Mtok
    $0.980 / $3.08
    chutes/fp8fp8
  • No data24h 99.69%
    7d
    99.44%
    30d
    99.44%
    p50
    577ms
    p99
    15.6s
    Throughput
    71 tps
    Context
    203K
    $/Mtok
    $1.20 / $4.40
    crusoe/fp8fp8
  • No data24h 99.69%
    7d
    99.40%
    30d
    99.40%
    p50
    1.45s
    p99
    17.7s
    Throughput
    41 tps
    Context
    203K
    $/Mtok
    $1.05 / $3.50
    deepinfra/fp4fp4
  • No data24h 99.86%
    7d
    99.34%
    30d
    99.34%
    p50
    2.93s
    p99
    10.5s
    Throughput
    36 tps
    Context
    205K
    $/Mtok
    $1.38 / $4.40
    novita/fp8fp8
  • No data24h 77.33%
    7d
    84.04%
    30d
    84.04%
    p50
    3.84s
    p99
    24.8s
    Throughput
    25 tps
    Context
    203K
    $/Mtok
    $1.21 / $4.20
    phala
  • Down24h 97.86%
    7d
    96.49%
    30d
    96.49%
    p50
    6.55s
    p99
    17.5s
    Throughput
    11 tps
    Context
    200K
    $/Mtok
    $1.54 / $4.84
    venice/fp8fp8

Measured 2m ago. Latency and throughput are OpenRouter's rolling 30-minute windows. ◆ marks the best value in each column among these endpoints.

Price vs latency

One dot per endpoint at the latest round · further left = faster, lower = cheaper · hover a dot for its tag

History

Availability

Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.

Time to first token

Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.

Throughput

Median tokens per second.