modelstatus.dev

Qwen3.5-122B-A10B

qwen/qwen3.5-122b-a10b

Compare with Qwen3.5 397B A17B, Qwen3.5-35B-A3B, Qwen3.5-27B

CompareQwen3.5-122B-A10B
  • Alibaba100%
    Degraded24h 99.73%
    p50
    386ms
    p99
    2.01s
    Throughput
    25 tps
    Context
    262K
    $/Mtok
    $0.260 / $2.08
    alibaba
  • DeepInfra100%
    Healthy24h 99.42%
    p50
    288ms
    p99
    2.47s
    Throughput
    37 tps
    Context
    262K
    $/Mtok
    $0.290 / $2.40
    deepinfra/fp4fp4
  • Novita100%
    Healthy24h 99.60%
    p50
    1.02s
    p99
    5.17s
    Throughput
    18 tps
    Context
    262K
    $/Mtok
    $0.400 / $3.20
    novita/bf16bf16
  • SiliconFlow100%
    Healthy24h 98.55%
    p50
    1.93s
    p99
    4.51s
    Throughput
    29 tps
    Context
    262K
    $/Mtok
    $0.260 / $2.08
    siliconflow/fp8fp8
  • AtlasCloud
    No data24h 99.52%
    p50
    903ms
    p99
    1.77s
    Throughput
    12 tps
    Context
    262K
    $/Mtok
    $0.300 / $2.40
    atlas-cloud/fp8fp8

Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.

1 endpoint reporting no data is omitted from the charts.

Availability

Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.

Time to first token

Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows, averaged over the hour.

Throughput

Median tokens per second.