Gemini 3.6 Flash
google/gemini-3.6-flash
Compare with Gemini 3.5 Flash, Gemini 3.7 Flash, Gemini 3.5 Flash Lite
CompareGemini 3.6 Flash
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
Google AI Studiogoogle-ai-studio | Healthy | 99.85% | 98.61% | 1.61s | 6.33s | 132 tps | 1.0M | $0.750 / $3.75 |
Google AI Studiogoogle-ai-studio/flex | Healthy | 99.85% | 98.61% | 1.96s | 24.2s | 158 tps | 1.0M | $0.375 / $1.88 |
Google AI Studiogoogle-ai-studio/priority | Healthy | 99.85% | 98.61% | 778ms | 843ms | 68 tps | 1.0M | $1.35 / $6.75 |
Googlegoogle-vertex/global | Healthy | 99.53% | 98.63% | 2.11s | 12.9s | 148 tps | 1.0M | $0.750 / $3.75 |
Googlegoogle-vertex/global/flex | Healthy | 99.53% | 98.63% | 7.86s | 11.6s | 78 tps | 1.0M | $0.375 / $1.88 |
Googlegoogle-vertex/global/priority | Healthy | 99.53% | 98.63% | — | — | — | 1.0M | $1.35 / $6.75 |
Googlegoogle-vertex/us | No data | — | 100% | — | — | — | 1.0M | $0.825 / $4.13 |
- Google AI Studio99.85%Healthy24h 98.61%
- p50
- 1.61s
- p99
- 6.33s
- Throughput
- 132 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-ai-studio - Google AI Studio99.85%Healthy24h 98.61%
- p50
- 1.96s
- p99
- 24.2s
- Throughput
- 158 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-ai-studio/flex - Google AI Studio99.85%Healthy24h 98.61%
- p50
- 778ms
- p99
- 843ms
- Throughput
- 68 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-ai-studio/priority - Google99.53%Healthy24h 98.63%
- p50
- 2.11s
- p99
- 12.9s
- Throughput
- 148 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-vertex/global - Google99.53%Healthy24h 98.63%
- p50
- 7.86s
- p99
- 11.6s
- Throughput
- 78 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-vertex/global/flex - Google99.53%Healthy24h 98.63%
- p50
- —
- p99
- —
- Throughput
- —
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-vertex/global/priority - Google—No data24h 100%
- p50
- —
- p99
- —
- Throughput
- —
- Context
- 1.0M
- $/Mtok
- $0.825 / $4.13
google-vertex/us
Measured 5m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
1 endpoint reporting no data is omitted from the charts.
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows.
Throughput
Median tokens per second.