Nemotron 3 Ultra
nvidia/nemotron-3-ultra-550b-a55b
Compare with Nemotron 3 Ultra (batch), Nemotron 3 Ultra (free), Nemotron 3 Nano 30B A3B
Showing only the venice/fp8 provider · show all 4
CompareNemotron 3 Ultra
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
| Healthy | 100% | 74.50% | 82.21% | 82.21% | 1.36s | 11.9s | 45 tps | 256K | $0.625 / $3.13 |
- Venice100%Healthy24h 74.50%
- 7d
- 82.21%
- 30d
- 82.21%
- p50
- 1.36s
- p99
- 11.9s
- Throughput
- 45 tps
- Context
- 256K
- $/Mtok
- $0.625 / $3.13
venice/fp8fp8
Measured 3m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.
Throughput
Median tokens per second.