GPT-5 Mini
openai/gpt-5-mini
Compare with GPT-5 Nano, GPT-5, GPT Audio
CompareGPT-5 Mini
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
Azureazure | Healthy | 100% | 100% | 2.13s | 18.0s | 47 tps | 400K | $0.250 / $2.00 |
OpenAIopenai | Healthy | 99.96% | 99.97% | 3.62s | 20.4s | 89 tps | 400K | $0.250 / $2.00 |
OpenAIopenai/flex | Healthy | 99.96% | 99.97% | 5.44s | 51.3s | 105 tps | 400K | $0.125 / $1.00 |
Azureazure/swedencentral | No data | — | 100% | 3.44s | 5.11s | 86 tps | 400K | $0.275 / $2.20 |
- Azure100%Healthy24h 100%
- p50
- 2.13s
- p99
- 18.0s
- Throughput
- 47 tps
- Context
- 400K
- $/Mtok
- $0.250 / $2.00
azure - OpenAI99.96%Healthy24h 99.97%
- p50
- 3.62s
- p99
- 20.4s
- Throughput
- 89 tps
- Context
- 400K
- $/Mtok
- $0.250 / $2.00
openai - OpenAI99.96%Healthy24h 99.97%
- p50
- 5.44s
- p99
- 51.3s
- Throughput
- 105 tps
- Context
- 400K
- $/Mtok
- $0.125 / $1.00
openai/flex - Azure—No data24h 100%
- p50
- 3.44s
- p99
- 5.11s
- Throughput
- 86 tps
- Context
- 400K
- $/Mtok
- $0.275 / $2.20
azure/swedencentral
Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
1 endpoint reporting no data is omitted from the charts.
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows.
Throughput
Median tokens per second.