gpt-oss-120b
openai/gpt-oss-120b
Compare with GPT Audio, GPT Audio Mini, GPT Chat Latest
Showing only the parasail/fp4 provider · show all 20
Comparegpt-oss-120b
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
| Healthy | 100% | 99.67% | 99.83% | 99.83% | 326ms | 1.72s | 98 tps | 131K | $0.100 / $0.750 |
- Parasail100%Healthy24h 99.67%
- 7d
- 99.83%
- 30d
- 99.83%
- p50
- 326ms
- p99
- 1.72s
- Throughput
- 98 tps
- Context
- 131K
- $/Mtok
- $0.100 / $0.750
parasail/fp4fp4
Measured 2m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
Price changes
- gpt-oss-120bMancer 2mancer/fp8$0.080 / $0.500 → $0.085 / $0.5002026-08-20 23:41 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.083 / $0.500 → $0.080 / $0.5002026-08-20 22:26 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.085 / $0.500 → $0.083 / $0.5002026-08-20 21:56 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.080 / $0.500 → $0.085 / $0.5002026-08-20 21:21 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.085 / $0.500 → $0.080 / $0.5002026-08-19 19:25 UTC
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window, averaged over the hour.
Throughput
Median tokens per second.