MiniMax M3
minimax/minimax-m3
Compare with MiniMax M2.7, MiniMax M2.5, MiniMax M3 (batch)
CompareMiniMax M3
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
AtlasCloudatlas-cloud/fp8fp8 | Healthy | 100% | 99.33% | 882ms | 4.07s | 87 tps | 524K | $0.300 / $1.20 |
Morphmorph/fp4fp4 | Healthy | 100% | 95.74% | 2.08s | 10.9s | 26 tps | 256K | $0.300 / $1.20 |
Venicevenice/fp8fp8 | Healthy | 100% | 99.83% | 863ms | 7.65s | 117 tps | 524K | $0.300 / $1.20 |
ModelRunmodelrun/fp4fp4 | Healthy | 99.84% | 99.46% | 1.83s | 7.07s | 123 tps | 1.0M | $0.750 / $3.00 |
Togethertogether | Healthy | 99.68% | 99.73% | 756ms | 10.7s | 81 tps | 524K | $0.300 / $1.20 |
Novitanovita/fp8fp8 | Healthy | 99.68% | 99.49% | 1.55s | 13.4s | 64 tps | 1.0M | $0.300 / $1.20 |
Minimaxminimax/fp8fp8 | Healthy | 99.63% | 99.59% | 939ms | 8.39s | 81 tps | 524K | $0.300 / $1.20 |
StreamLakestreamlake/fp8fp8 | Healthy | 99.08% | 98.58% | 1.71s | 13.0s | 96 tps | 1.0M | $0.300 / $1.20 |
GMICloudgmicloud/fp8fp8 | Healthy | 99.03% | 99.12% | 3.07s | 8.87s | 42 tps | 1.0M | $0.240 / $0.960 |
CoreWeavecoreweave/fp4fp4 | Healthy | 98.71% | 98.51% | 667ms | 12.8s | 64 tps | 262K | $0.230 / $0.960 |
Parasailparasail/fp8fp8 | Healthy | 98.63% | 97.95% | 695ms | 8.49s | 140 tps | 1.0M | $0.300 / $1.20 |
DeepInfradeepinfra/fp8fp8 | Healthy | 97.78% | 98.10% | 816ms | 53.1s | 28 tps | 524K | $0.280 / $1.10 |
- AtlasCloud100%Healthy24h 99.33%
- p50
- 882ms
- p99
- 4.07s
- Throughput
- 87 tps
- Context
- 524K
- $/Mtok
- $0.300 / $1.20
atlas-cloud/fp8fp8 - Morph100%Healthy24h 95.74%
- p50
- 2.08s
- p99
- 10.9s
- Throughput
- 26 tps
- Context
- 256K
- $/Mtok
- $0.300 / $1.20
morph/fp4fp4 - Venice100%Healthy24h 99.83%
- p50
- 863ms
- p99
- 7.65s
- Throughput
- 117 tps
- Context
- 524K
- $/Mtok
- $0.300 / $1.20
venice/fp8fp8 - ModelRun99.84%Healthy24h 99.46%
- p50
- 1.83s
- p99
- 7.07s
- Throughput
- 123 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.00
modelrun/fp4fp4 - Together99.68%Healthy24h 99.73%
- p50
- 756ms
- p99
- 10.7s
- Throughput
- 81 tps
- Context
- 524K
- $/Mtok
- $0.300 / $1.20
together - Novita99.68%Healthy24h 99.49%
- p50
- 1.55s
- p99
- 13.4s
- Throughput
- 64 tps
- Context
- 1.0M
- $/Mtok
- $0.300 / $1.20
novita/fp8fp8 - Minimax99.63%Healthy24h 99.59%
- p50
- 939ms
- p99
- 8.39s
- Throughput
- 81 tps
- Context
- 524K
- $/Mtok
- $0.300 / $1.20
minimax/fp8fp8 - StreamLake99.08%Healthy24h 98.58%
- p50
- 1.71s
- p99
- 13.0s
- Throughput
- 96 tps
- Context
- 1.0M
- $/Mtok
- $0.300 / $1.20
streamlake/fp8fp8 - GMICloud99.03%Healthy24h 99.12%
- p50
- 3.07s
- p99
- 8.87s
- Throughput
- 42 tps
- Context
- 1.0M
- $/Mtok
- $0.240 / $0.960
gmicloud/fp8fp8 - CoreWeave98.71%Healthy24h 98.51%
- p50
- 667ms
- p99
- 12.8s
- Throughput
- 64 tps
- Context
- 262K
- $/Mtok
- $0.230 / $0.960
coreweave/fp4fp4 - Parasail98.63%Healthy24h 97.95%
- p50
- 695ms
- p99
- 8.49s
- Throughput
- 140 tps
- Context
- 1.0M
- $/Mtok
- $0.300 / $1.20
parasail/fp8fp8 - DeepInfra97.78%Healthy24h 98.10%
- p50
- 816ms
- p99
- 53.1s
- Throughput
- 28 tps
- Context
- 524K
- $/Mtok
- $0.280 / $1.10
deepinfra/fp8fp8
Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows, averaged over the hour.
Throughput
Median tokens per second.