gpt-oss-120b
openai/gpt-oss-120b
Compare with GPT Audio, GPT Audio Mini, GPT Chat Latest
Showing only the mancer/fp8 provider · show all 20
Comparegpt-oss-120b
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
| Healthy | 100% | 99.36% | 99.75% | 99.75% | 640ms | 3.49s | 103 tps | 131K | $0.085 / $0.500 |
- Mancer 2100%Healthy24h 99.36%
- 7d
- 99.75%
- 30d
- 99.75%
- p50
- 640ms
- p99
- 3.49s
- Throughput
- 103 tps
- Context
- 131K
- $/Mtok
- $0.085 / $0.500
mancer/fp8fp8
Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
Price changes
- gpt-oss-120bMancer 2mancer/fp8$0.080 / $0.500 → $0.085 / $0.5002026-08-20 23:41 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.083 / $0.500 → $0.080 / $0.5002026-08-20 22:26 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.085 / $0.500 → $0.083 / $0.5002026-08-20 21:56 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.080 / $0.500 → $0.085 / $0.5002026-08-20 21:21 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.085 / $0.500 → $0.080 / $0.5002026-08-19 19:25 UTC
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.
Throughput
Median tokens per second.