modelstatus.dev

Kimi K2.6

Kimi K2.6: providers compared

Every endpoint serving this model, side by side. ◆ marks the best value in each column — per model, never fleet-wide.

  • Cloudflare100% (best)
    Healthy24h 99.98%
    7d
    87.44%
    30d
    87.44%
    p50
    700ms
    p99
    5.48s
    Throughput
    41 tps
    Context
    262K
    $/Mtok
    $0.950 / $4.00
    cloudflare
  • Inceptron100% (best)
    Healthy24h 99.68%
    7d
    99.45%
    30d
    99.45%
    p50
    766ms
    p99
    7.82s
    Throughput
    48 tps
    Context
    262K
    $/Mtok
    $0.600 / $3.41
    inceptron/int4int4
  • Moonshot AI100% (best)
    Healthy24h 99.97%
    7d
    99.91% (best)
    30d
    99.91% (best)
    p50
    2.26s
    p99
    32.4s
    Throughput
    31 tps
    Context
    262K
    $/Mtok
    $0.950 / $4.00
    moonshotai/int4int4
  • Novita100% (best)
    Healthy24h 99.75%
    7d
    99.67%
    30d
    99.67%
    p50
    1.33s
    p99
    16.4s
    Throughput
    27 tps
    Context
    262K
    $/Mtok
    $0.800 / $3.40
    novita
  • Parasail100% (best)
    Healthy24h 99.61%
    7d
    99.11%
    30d
    99.11%
    p50
    925ms
    p99
    10.5s
    Throughput
    76 tps
    Context
    262K
    $/Mtok
    $0.750 / $3.50
    parasail/int4int4
  • SiliconFlow100% (best)
    Healthy24h 99.82%
    7d
    99.81%
    30d
    99.81%
    p50
    1.13s
    p99
    8.54s
    Throughput
    44 tps
    Context
    262K
    $/Mtok
    $0.770 / $3.40
    siliconflow/fp8fp8
  • StreamLake100% (best)
    Healthy24h 97.88%
    7d
    98.23%
    30d
    98.23%
    p50
    1.69s
    p99
    23.4s
    Throughput
    38 tps
    Context
    256K
    $/Mtok
    $0.598 / $2.52
    streamlake/fp8fp8
  • Venice100% (best)
    Healthy24h 96.16%
    7d
    96.52%
    30d
    96.52%
    p50
    1.33s
    p99
    10.4s
    Throughput
    17 tps
    Context
    256K
    $/Mtok
    $0.750 / $3.50
    venice/int4int4
  • Decart99.51%
    Healthy24h 98.12%
    7d
    97.54%
    30d
    97.54%
    p50
    664ms
    p99
    8.56s
    Throughput
    64 tps
    Context
    262K
    $/Mtok
    $0.568 / $2.39 (best)
    decart/fp4fp4
  • Baidu99.43%
    Healthy24h 99.84%
    7d
    99.82%
    30d
    99.82%
    p50
    1.67s
    p99
    10.2s
    Throughput
    50 tps
    Context
    262K
    $/Mtok
    $0.580 / $2.44
    baidu/fp4fp4
  • DeepInfra98.06%
    Healthy24h 99.73%
    7d
    99.15%
    30d
    99.15%
    p50
    1.07s
    p99
    8.24s
    Throughput
    15 tps
    Context
    262K
    $/Mtok
    $0.750 / $3.50
    deepinfra/fp4fp4
  • CoreWeave97.75%
    Healthy24h 99.33%
    7d
    99.42%
    30d
    99.42%
    p50
    949ms
    p99
    12.6s
    Throughput
    123 tps (best)
    Context
    262K
    $/Mtok
    $0.650 / $3.41
    coreweave/fp4fp4
  • Chutes95.16%
    Healthy24h 98.17%
    7d
    98.35%
    30d
    98.35%
    p50
    2.31s
    p99
    38.9s
    Throughput
    39 tps
    Context
    262K
    $/Mtok
    $0.580 / $3.40
    chutes/int4int4
  • No data24h 95.96%
    7d
    96.19%
    30d
    96.19%
    p50
    2.48s
    p99
    15.5s
    Throughput
    29 tps
    Context
    262K
    $/Mtok
    $0.950 / $4.00
    atlas-cloud/int4int4
  • No data24h
    7d
    30d
    p50
    852ms
    p99
    4.27s (best)
    Throughput
    81 tps
    Context
    262K
    $/Mtok
    $0.950 / $4.00
    baseten/fp4fp4
  • Down24h 90.79%
    7d
    88.17%
    30d
    88.17%
    p50
    813ms
    p99
    112s
    Throughput
    8.0 tps
    Context
    262K
    $/Mtok
    $0.700 / $3.50
    crusoe/bf16bf16
  • Down24h 98.92%
    7d
    98.98%
    30d
    98.98%
    p50
    2.57s
    p99
    51.5s
    Throughput
    15 tps
    Context
    262K
    $/Mtok
    $0.950 / $4.00
    digitalocean
  • No data24h 82.26%
    7d
    96.85%
    30d
    96.85%
    p50
    480ms (best)
    p99
    6.02s
    Throughput
    75 tps
    Context
    262K
    $/Mtok
    $0.950 / $4.00
    fireworks
  • No data24h 98.58%
    7d
    98.28%
    30d
    98.28%
    p50
    2.06s
    p99
    12.9s
    Throughput
    19 tps
    Context
    262K
    $/Mtok
    $1.09 / $4.60
    phala
  • No data24h
    7d
    99.88%
    30d
    99.88%
    p50
    p99
    Throughput
    Context
    262K
    $/Mtok
    $1.00 / $4.00
    sail-research/int4int4
  • No data24h
    7d
    97.92%
    30d
    97.92%
    p50
    p99
    Throughput
    Context
    262K
    $/Mtok
    $1.20 / $4.50
    together

Measured 4m ago. Latency and throughput are OpenRouter's rolling 30-minute windows. ◆ marks the best value in each column among these endpoints.

Price vs latency

One dot per endpoint at the latest round · further left = faster, lower = cheaper · hover a dot for its tag

History

Availability

Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.

Time to first token

Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.

Throughput

Median tokens per second.