Gemini 3.7 Flash
google/gemini-3.7-flash
Compare with Gemini 3.6 Flash, Gemini 3.5 Flash, Gemini 3.5 Flash Lite
CompareGemini 3.7 Flash
| Provider | State | Uptime 5m | 24h | p50 latency | p99 | Throughput | Context | $/Mtok |
|---|---|---|---|---|---|---|---|---|
Google AI Studiogoogle-ai-studio | Healthy | 99.72% | 99.58% | 1.99s | 11.6s | 183 tps | 1.0M | $0.750 / $3.75 |
Google AI Studiogoogle-ai-studio/flex | Healthy | 99.72% | 99.58% | 1.74s | 11.3s | 229 tps | 1.0M | $0.375 / $1.88 |
Google AI Studiogoogle-ai-studio/priority | Healthy | 99.72% | 99.58% | 2.55s | 6.48s | 101 tps | 1.0M | $1.35 / $6.75 |
Googlegoogle-vertex/global | Healthy | 99.53% | 98.72% | 2.43s | 15.3s | 84 tps | 1.0M | $0.375 / $1.88 |
Googlegoogle-vertex/global/flex | Healthy | 99.53% | 98.72% | 17.0s | 43.3s | 57 tps | 1.0M | $0.188 / $0.938 |
Googlegoogle-vertex/global/priority | Healthy | 99.53% | 98.72% | 2.14s | 6.21s | 115 tps | 1.0M | $0.675 / $3.38 |
- Google AI Studio99.72%Healthy24h 99.58%
- p50
- 1.99s
- p99
- 11.6s
- Throughput
- 183 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-ai-studio - Google AI Studio99.72%Healthy24h 99.58%
- p50
- 1.74s
- p99
- 11.3s
- Throughput
- 229 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-ai-studio/flex - Google AI Studio99.72%Healthy24h 99.58%
- p50
- 2.55s
- p99
- 6.48s
- Throughput
- 101 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-ai-studio/priority - Google99.53%Healthy24h 98.72%
- p50
- 2.43s
- p99
- 15.3s
- Throughput
- 84 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-vertex/global - Google99.53%Healthy24h 98.72%
- p50
- 17.0s
- p99
- 43.3s
- Throughput
- 57 tps
- Context
- 1.0M
- $/Mtok
- $0.188 / $0.938
google-vertex/global/flex - Google99.53%Healthy24h 98.72%
- p50
- 2.14s
- p99
- 6.21s
- Throughput
- 115 tps
- Context
- 1.0M
- $/Mtok
- $0.675 / $3.38
google-vertex/global/priority
Measured 5m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. p50 and p99 are OpenRouter's rolling 30-minute windows.
Throughput
Median tokens per second.