Gemini 3.6 Flash
google/gemini-3.6-flash
Compare with Gemini 3.5 Flash, Gemini 3.7 Flash, Gemini 3.5 Flash Lite
CompareGemini 3.6 Flash
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
Google AI Studiogoogle-ai-studio | Healthy | 98.62% | 98.52% | 1.76s | 26.2s | 125 tps | 1.0M | $0.750 / $3.75 |
Google AI Studiogoogle-ai-studio/flex | Healthy | 98.62% | 98.52% | 1.66s | 16.3s | 143 tps | 1.0M | $0.375 / $1.88 |
Google AI Studiogoogle-ai-studio/priority | Healthy | 98.62% | 98.52% | 774ms | 966ms | 63 tps | 1.0M | $1.35 / $6.75 |
Googlegoogle-vertex/global | Healthy | 96.47% | 98.64% | 1.88s | 20.3s | 114 tps | 1.0M | $0.750 / $3.75 |
Googlegoogle-vertex/global/flex | Healthy | 96.47% | 98.64% | 7.73s | 49.2s | 91 tps | 1.0M | $0.375 / $1.88 |
Googlegoogle-vertex/global/priority | Healthy | 96.47% | 98.64% | 1.37s | 1.72s | 13 tps | 1.0M | $1.35 / $6.75 |
Googlegoogle-vertex/us | No data | — | 100% | 1.19s | 1.22s | 206 tps | 1.0M | $0.825 / $4.13 |
- Google AI Studio98.62%Healthy24h 98.52%
- p50
- 1.76s
- p99
- 26.2s
- Throughput
- 125 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-ai-studio - Google AI Studio98.62%Healthy24h 98.52%
- p50
- 1.66s
- p99
- 16.3s
- Throughput
- 143 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-ai-studio/flex - Google AI Studio98.62%Healthy24h 98.52%
- p50
- 774ms
- p99
- 966ms
- Throughput
- 63 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-ai-studio/priority - Google96.47%Healthy24h 98.64%
- p50
- 1.88s
- p99
- 20.3s
- Throughput
- 114 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-vertex/global - Google96.47%Healthy24h 98.64%
- p50
- 7.73s
- p99
- 49.2s
- Throughput
- 91 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-vertex/global/flex - Google96.47%Healthy24h 98.64%
- p50
- 1.37s
- p99
- 1.72s
- Throughput
- 13 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-vertex/global/priority - Google—No data24h 100%
- p50
- 1.19s
- p99
- 1.22s
- Throughput
- 206 tps
- Context
- 1.0M
- $/Mtok
- $0.825 / $4.13
google-vertex/us
Measured 1m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
1 endpoint reporting no data is omitted from the charts.
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows, averaged over the hour.
Throughput
Median tokens per second.