Gemini 3.7 Flash
google/gemini-3.7-flashInputOutput
Compare its 6 providersIncident RSSCompare with Gemini 3.8 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash
CompareGemini 3.7 Flash
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
Google AI Studiogoogle-ai-studio/flex | Healthy | 100% | 99.94% | 99.88% | 99.68% | 1.79s | 6.76s | 152 tps | 1.0M | $0.375 / $1.88 |
Google AI Studiogoogle-ai-studio | Healthy | 99.68% | 99.80% | 99.60% | 99.51% | 2.34s | 8.64s | 182 tps | 1.0M | $0.750 / $3.75 |
Googlegoogle-vertex/global | Healthy | 99.15% | 99.37% | 99.08% | 99.39% | 2.49s | 10.9s | 93 tps | 1.0M | $0.750 / $3.75 |
Googlegoogle-vertex/global/flex | No data | — | 98.64% | 99.57% | 81.03% | 13.0s | 211s | 21 tps | 1.0M | $0.375 / $1.88 |
Googlegoogle-vertex/global/priority | No data | — | 99.80% | 99.80% | 99.75% | 1.95s | 6.31s | 89 tps | 1.0M | $1.35 / $6.75 |
Google AI Studiogoogle-ai-studio/priority | No data | — | 99.94% | 99.86% | 99.61% | 1.59s | 4.19s | 94 tps | 1.0M | $1.35 / $6.75 |
- Google AI Studio100%Healthy24h 99.94%
- 7d
- 99.88%
- 30d
- 99.68%
- p50
- 1.79s
- p99
- 6.76s
- Throughput
- 152 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-ai-studio/flex - Google AI Studio99.68%Healthy24h 99.80%
- 7d
- 99.60%
- 30d
- 99.51%
- p50
- 2.34s
- p99
- 8.64s
- Throughput
- 182 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-ai-studio - Google99.15%Healthy24h 99.37%
- 7d
- 99.08%
- 30d
- 99.39%
- p50
- 2.49s
- p99
- 10.9s
- Throughput
- 93 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-vertex/global - No data24h 98.64%
- 7d
- 99.57%
- 30d
- 81.03%
- p50
- 13.0s
- p99
- 211s
- Throughput
- 21 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-vertex/global/flex - No data24h 99.80%
- 7d
- 99.80%
- 30d
- 99.75%
- p50
- 1.95s
- p99
- 6.31s
- Throughput
- 89 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-vertex/global/priority - No data24h 99.94%
- 7d
- 99.86%
- 30d
- 99.61%
- p50
- 1.59s
- p99
- 4.19s
- Throughput
- 94 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-ai-studio/priority
Measured 5m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window, averaged over the hour.
Throughput
Median tokens per second.
Daily availability
One row per endpoint · one dot per day · last 90 days.
worst hour ≥ 95%90–95%< 90%no data
Colour by each endpoint's worst hour of each UTC day · 6 endpoints · 47 days with data, 31 endpoint-days below 90%.
Price changes
- Gemini 3.7 FlashGooglegoogle-vertex/global$0.375 → $0.750 / $1.88 → $3.75cache read $0.037 → $0.075cache write $0.021 → $0.042reasoning $1.88 → $3.752026-08-28 18:10 UTC
- Gemini 3.7 FlashGooglegoogle-vertex/global/flex$0.188 → $0.375 / $0.938 → $1.88cache read $0.019 → $0.037cache write $0.010 → $0.021reasoning $0.938 → $1.882026-08-28 18:10 UTC
- Gemini 3.7 FlashGooglegoogle-vertex/global/priority$0.675 → $1.35 / $3.38 → $6.75cache read $0.068 → $0.135cache write $0.037 → $0.075reasoning $3.38 → $6.752026-08-28 18:10 UTC