Gemini 3.5 Flash Lite
google/gemini-3.5-flash-lite
Compare with Gemini 3.5 Flash, Gemini 3.6 Flash, Gemini 3.7 Flash
CompareGemini 3.5 Flash Lite
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
Googlegoogle-vertex/global | Healthy | 99.89% | 99.89% | 615ms | 5.54s | 37 tps | 1.0M | $0.300 / $2.50 |
Googlegoogle-vertex/global/flex | Healthy | 99.89% | 99.89% | 7.58s | 15.3s | 109 tps | 1.0M | $0.150 / $1.25 |
Googlegoogle-vertex/global/priority | Healthy | 99.89% | 99.89% | 4.01s | 6.40s | 93 tps | 1.0M | $0.540 / $4.50 |
Google AI Studiogoogle-ai-studio | Healthy | 99.89% | 99.87% | 538ms | 5.13s | 54 tps | 1.0M | $0.300 / $2.50 |
Google AI Studiogoogle-ai-studio/flex | Healthy | 99.89% | 99.87% | 679ms | 3.45s | 27 tps | 1.0M | $0.150 / $1.25 |
Google AI Studiogoogle-ai-studio/priority | Healthy | 99.89% | 99.87% | 426ms | 1.11s | 91 tps | 1.0M | $0.540 / $4.50 |
Googlegoogle-vertex/us | No data | — | — | — | — | — | 1.0M | $0.330 / $2.75 |
- Google99.89%Healthy24h 99.89%
- p50
- 615ms
- p99
- 5.54s
- Throughput
- 37 tps
- Context
- 1.0M
- $/Mtok
- $0.300 / $2.50
google-vertex/global - Google99.89%Healthy24h 99.89%
- p50
- 7.58s
- p99
- 15.3s
- Throughput
- 109 tps
- Context
- 1.0M
- $/Mtok
- $0.150 / $1.25
google-vertex/global/flex - Google99.89%Healthy24h 99.89%
- p50
- 4.01s
- p99
- 6.40s
- Throughput
- 93 tps
- Context
- 1.0M
- $/Mtok
- $0.540 / $4.50
google-vertex/global/priority - Google AI Studio99.89%Healthy24h 99.87%
- p50
- 538ms
- p99
- 5.13s
- Throughput
- 54 tps
- Context
- 1.0M
- $/Mtok
- $0.300 / $2.50
google-ai-studio - Google AI Studio99.89%Healthy24h 99.87%
- p50
- 679ms
- p99
- 3.45s
- Throughput
- 27 tps
- Context
- 1.0M
- $/Mtok
- $0.150 / $1.25
google-ai-studio/flex - Google AI Studio99.89%Healthy24h 99.87%
- p50
- 426ms
- p99
- 1.11s
- Throughput
- 91 tps
- Context
- 1.0M
- $/Mtok
- $0.540 / $4.50
google-ai-studio/priority - Google—No data24h —
- p50
- —
- p99
- —
- Throughput
- —
- Context
- 1.0M
- $/Mtok
- $0.330 / $2.75
google-vertex/us
Measured 5m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
1 endpoint reporting no data is omitted from the charts.
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows.
Throughput
Median tokens per second.