gpt-oss-120b
openai/gpt-oss-120b
Compare with GPT Audio, GPT Audio Mini, GPT Chat Latest
Showing only the mancer/fp8 provider · show all 20
Comparegpt-oss-120b
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
| Healthy | 99.17% | 99.26% | 99.76% | 99.76% | 842ms | 9.08s | 28 tps | 131K | $0.085 / $0.500 |
- Mancer 299.17%Healthy24h 99.26%
- 7d
- 99.76%
- 30d
- 99.76%
- p50
- 842ms
- p99
- 9.08s
- Throughput
- 28 tps
- Context
- 131K
- $/Mtok
- $0.085 / $0.500
mancer/fp8fp8
Measured 4m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
Price changes
- gpt-oss-120bMancer 2mancer/fp8$0.080 / $0.500 → $0.085 / $0.5002026-08-20 23:41 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.083 / $0.500 → $0.080 / $0.5002026-08-20 22:26 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.085 / $0.500 → $0.083 / $0.5002026-08-20 21:56 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.080 / $0.500 → $0.085 / $0.5002026-08-20 21:21 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.085 / $0.500 → $0.080 / $0.5002026-08-19 19:25 UTC
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.
Throughput
Median tokens per second.