gpt-oss-120b
openai/gpt-oss-120b
Compare with GPT Audio, GPT Audio Mini, GPT Chat Latest
Compare its 20 providers · Incident RSS
Showing only the mara provider · show all 20
Comparegpt-oss-120b
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
Maramara | Degraded | 91.63% | 97.40% | 95.46% | 95.46% | 3.07s | 11.9s | 104 tps | 131K | $0.150 / $0.750 |
- Mara91.63%Degraded24h 97.40%
- 7d
- 95.46%
- 30d
- 95.46%
- p50
- 3.07s
- p99
- 11.9s
- Throughput
- 104 tps
- Context
- 131K
- $/Mtok
- $0.150 / $0.750
mara
Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
Price history
Input $/Mtok
Step line — each recorded price holds until the next change.
Output $/Mtok
Step line — each recorded price holds until the next change.
Price changes
- gpt-oss-120bMancer 2mancer/fp8$0.080 → $0.085 / $0.5002026-08-20 23:41 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.083 → $0.080 / $0.5002026-08-20 22:26 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.085 → $0.083 / $0.5002026-08-20 21:56 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.080 → $0.085 / $0.5002026-08-20 21:21 UTC
- gpt-oss-120bMancer 2mancer/fp8$0.085 → $0.080 / $0.5002026-08-19 19:25 UTC
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.
Throughput
Median tokens per second.
Daily availability
Half a year, one cell per UTC day.
Feb
Mar
Apr
May
Jun
Jul
Aug
Mon
Wed
Fri
worst hour ≥ 95%90–95%< 90%no data
Colour by worst endpoint-hour of each UTC day · 3 days with data, 3 with at least one endpoint-hour below 90%.