Gemini 3.1 Flash Lite
google/gemini-3.1-flash-liteInputOutput
Compare its 8 providersIncident RSSCompare with Gemini 3.5 Flash, Gemini 3.6 Flash, Gemini 2.5 Flash
CompareGemini 3.1 Flash Lite
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
Googlegoogle-vertex/eu | Healthy | 100% | 100% | 99.98% | 99.97% | 650ms | 9.66s | 187 tps | 1.0M | $0.275 / $1.65 |
Googlegoogle-vertex/global/flex | Healthy | 100% | 99.94% | 99.95% | 57.57% | 16.0s | 26.4s | 4.0 tps | 1.0M | $0.125 / $0.750 |
Googlegoogle-vertex/global/priority | Healthy | 100% | 100% | 99.99% | 99.96% | 1.20s | 2.53s | 96 tps | 1.0M | $0.450 / $2.70 |
Google AI Studiogoogle-ai-studio/flex | Healthy | 100% | 99.99% | 99.93% | 99.95% | 478ms | 5.30s | 217 tps | 1.0M | $0.125 / $0.750 |
Google AI Studiogoogle-ai-studio/priority | Healthy | 100% | 99.99% | 99.94% | 99.93% | 1.07s | 3.75s | 159 tps | 1.0M | $0.450 / $2.70 |
Googlegoogle-vertex/global | Healthy | 99.98% | 99.96% | 99.49% | 99.72% | 943ms | 4.34s | 99 tps | 1.0M | $0.250 / $1.50 |
Google AI Studiogoogle-ai-studio | Healthy | 99.91% | 99.93% | 99.82% | 99.81% | 720ms | 3.39s | 120 tps | 1.0M | $0.250 / $1.50 |
Googlegoogle-vertex/us | No data | — | 100% | 99.99% | 99.96% | 1.09s | 1.92s | 276 tps | 1.0M | $0.275 / $1.65 |
- Google100%Healthy24h 100%
- 7d
- 99.98%
- 30d
- 99.97%
- p50
- 650ms
- p99
- 9.66s
- Throughput
- 187 tps
- Context
- 1.0M
- $/Mtok
- $0.275 / $1.65
google-vertex/eu - Google100%Healthy24h 99.94%
- 7d
- 99.95%
- 30d
- 57.57%
- p50
- 16.0s
- p99
- 26.4s
- Throughput
- 4.0 tps
- Context
- 1.0M
- $/Mtok
- $0.125 / $0.750
google-vertex/global/flex - Google100%Healthy24h 100%
- 7d
- 99.99%
- 30d
- 99.96%
- p50
- 1.20s
- p99
- 2.53s
- Throughput
- 96 tps
- Context
- 1.0M
- $/Mtok
- $0.450 / $2.70
google-vertex/global/priority - Google AI Studio100%Healthy24h 99.99%
- 7d
- 99.93%
- 30d
- 99.95%
- p50
- 478ms
- p99
- 5.30s
- Throughput
- 217 tps
- Context
- 1.0M
- $/Mtok
- $0.125 / $0.750
google-ai-studio/flex - Google AI Studio100%Healthy24h 99.99%
- 7d
- 99.94%
- 30d
- 99.93%
- p50
- 1.07s
- p99
- 3.75s
- Throughput
- 159 tps
- Context
- 1.0M
- $/Mtok
- $0.450 / $2.70
google-ai-studio/priority - Google99.98%Healthy24h 99.96%
- 7d
- 99.49%
- 30d
- 99.72%
- p50
- 943ms
- p99
- 4.34s
- Throughput
- 99 tps
- Context
- 1.0M
- $/Mtok
- $0.250 / $1.50
google-vertex/global - Google AI Studio99.91%Healthy24h 99.93%
- 7d
- 99.82%
- 30d
- 99.81%
- p50
- 720ms
- p99
- 3.39s
- Throughput
- 120 tps
- Context
- 1.0M
- $/Mtok
- $0.250 / $1.50
google-ai-studio - No data24h 100%
- 7d
- 99.99%
- 30d
- 99.96%
- p50
- 1.09s
- p99
- 1.92s
- Throughput
- 276 tps
- Context
- 1.0M
- $/Mtok
- $0.275 / $1.65
google-vertex/us
Measured 2m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.
Throughput
Median tokens per second.
Daily availability
One row per endpoint · one dot per day · last 90 days.
worst hour ≥ 95%90–95%< 90%no data
Colour by each endpoint's worst hour of each UTC day · 8 endpoints · 47 days with data, 25 endpoint-days below 90%.