Llama 3 8B Lunaris
sao10k/l3-lunaris-8b
Compare with Llama 3.1 Euryale 70B v2.2, Llama 3.3 Euryale 70B, GLM 5.2
CompareLlama 3 8B Lunaris
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
DeepInfradeepinfra/turbofp8 | Healthy | 100% | 98.92% | 143ms | 977ms | 71 tps | 8K | $0.040 / $0.050 |
Novitanovita/bf16bf16 | Healthy | 100% | 99.00% | 336ms | 1.12s | 67 tps | 8K | $0.050 / $0.050 |
Parasailparasail/bf16bf16 | Healthy | 100% | 99.63% | 499ms | 1.47s | 64 tps | 8K | $0.040 / $0.050 |
- DeepInfra100%Healthy24h 98.92%
- p50
- 143ms
- p99
- 977ms
- Throughput
- 71 tps
- Context
- 8K
- $/Mtok
- $0.040 / $0.050
deepinfra/turbofp8 - Novita100%Healthy24h 99.00%
- p50
- 336ms
- p99
- 1.12s
- Throughput
- 67 tps
- Context
- 8K
- $/Mtok
- $0.050 / $0.050
novita/bf16bf16 - Parasail100%Healthy24h 99.63%
- p50
- 499ms
- p99
- 1.47s
- Throughput
- 64 tps
- Context
- 8K
- $/Mtok
- $0.040 / $0.050
parasail/bf16bf16
Measured 2m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows.
Throughput
Median tokens per second.