Gemini 2.5 Flash Lite
google/gemini-2.5-flash-lite
Compare with Gemini 2.5 Flash, Gemini 2.5 Flash (batch), Gemini 2.5 Pro
CompareGemini 2.5 Flash Lite
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
Googlegoogle-vertex/eu | Healthy | 99.86% | 99.92% | 395ms | 3.38s | 105 tps | 1.0M | $0.100 / $0.400 |
Googlegoogle-vertex | Healthy | 99.78% | 99.82% | 422ms | 2.93s | 81 tps | 1.0M | $0.100 / $0.400 |
Google AI Studiogoogle-ai-studio | Healthy | 99.21% | 99.41% | 457ms | 7.39s | 169 tps | 1.0M | $0.100 / $0.400 |
Google AI Studiogoogle-ai-studio/flex | Healthy | 99.21% | 99.41% | 937ms | 15.4s | 53 tps | 1.0M | $0.050 / $0.200 |
Google AI Studiogoogle-ai-studio/priority | Healthy | 99.21% | 99.41% | 324ms | 379ms | 16 tps | 1.0M | $0.180 / $0.720 |
- Google99.86%Healthy24h 99.92%
- p50
- 395ms
- p99
- 3.38s
- Throughput
- 105 tps
- Context
- 1.0M
- $/Mtok
- $0.100 / $0.400
google-vertex/eu - Google99.78%Healthy24h 99.82%
- p50
- 422ms
- p99
- 2.93s
- Throughput
- 81 tps
- Context
- 1.0M
- $/Mtok
- $0.100 / $0.400
google-vertex - Google AI Studio99.21%Healthy24h 99.41%
- p50
- 457ms
- p99
- 7.39s
- Throughput
- 169 tps
- Context
- 1.0M
- $/Mtok
- $0.100 / $0.400
google-ai-studio - Google AI Studio99.21%Healthy24h 99.41%
- p50
- 937ms
- p99
- 15.4s
- Throughput
- 53 tps
- Context
- 1.0M
- $/Mtok
- $0.050 / $0.200
google-ai-studio/flex - Google AI Studio99.21%Healthy24h 99.41%
- p50
- 324ms
- p99
- 379ms
- Throughput
- 16 tps
- Context
- 1.0M
- $/Mtok
- $0.180 / $0.720
google-ai-studio/priority
Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows.
Throughput
Median tokens per second.