Llama 3.3 70B Instruct
meta-llama/llama-3.3-70b-instructInputOutput
Compare its 14 providersIncident RSSCompare with Llama 3.2 3B Instruct, Llama 3.2 1B Instruct, Llama 3.1 8B Instruct
Showing only the google-vertex/us-central1 provider · show all 14
CompareLlama 3.3 70B Instruct
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
Googlegoogle-vertex/us-central1 | No data | — | — | — | 99.99% | 1.10s | 1.10s | 115 tps | 128K | $0.720 / $0.720 |
- No data24h —
- 7d
- —
- 30d
- 99.99%
- p50
- 1.10s
- p99
- 1.10s
- Throughput
- 115 tps
- Context
- 128K
- $/Mtok
- $0.720 / $0.720
google-vertex/us-central1
Measured 3m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.
Throughput
Median tokens per second.
Daily availability
One row per endpoint · one dot per day · last 90 days.
worst hour ≥ 95%90–95%< 90%no data
Colour by each endpoint's worst hour of each UTC day · 14 endpoints · 49 days with data, 109 endpoint-days below 90%. Tables and charts above are filtered to google-vertex/us-central1.
Price changes
- Llama 3.3 70B InstructAkashMLakashml/fp8$0.130 → $0.200 / $0.400 → $0.520cache read — → $0.1002026-08-19 09:50 UTC