GPT-5 Nano
openai/gpt-5-nano
Compare with GPT-5.4 Nano, GPT-5 Nano (batch), GPT-4.1 Nano
CompareGPT-5 Nano
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
Azureazure | Healthy | 100% | 100% | 3.19s | 25.1s | 66 tps | 400K | $0.050 / $0.400 |
OpenAIopenai | Healthy | 96.93% | 97.82% | 644ms | 7.95s | 59 tps | 400K | $0.050 / $0.400 |
OpenAIopenai/flex | Healthy | 96.93% | 97.82% | 1.05s | 4.67s | 65 tps | 400K | $0.025 / $0.200 |
Azureazure/swedencentral | No data | — | 100% | — | — | — | 400K | $0.055 / $0.440 |
- Azure100%Healthy24h 100%
- p50
- 3.19s
- p99
- 25.1s
- Throughput
- 66 tps
- Context
- 400K
- $/Mtok
- $0.050 / $0.400
azure - OpenAI96.93%Healthy24h 97.82%
- p50
- 644ms
- p99
- 7.95s
- Throughput
- 59 tps
- Context
- 400K
- $/Mtok
- $0.050 / $0.400
openai - OpenAI96.93%Healthy24h 97.82%
- p50
- 1.05s
- p99
- 4.67s
- Throughput
- 65 tps
- Context
- 400K
- $/Mtok
- $0.025 / $0.200
openai/flex - Azure—No data24h 100%
- p50
- —
- p99
- —
- Throughput
- —
- Context
- 400K
- $/Mtok
- $0.055 / $0.440
azure/swedencentral
Measured 5m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
1 endpoint reporting no data is omitted from the charts.
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows.
Throughput
Median tokens per second.