Gemini 3.7 Flash
google/gemini-3.7-flashInputOutput
Compare its 6 providersIncident RSSCompare with Gemini 3.8 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash
CompareGemini 3.7 Flash
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
Google AI Studiogoogle-ai-studio | Healthy | 100% | 99.88% | 99.60% | 99.51% | 2.38s | 10.2s | 135 tps | 1.0M | $0.750 / $3.75 |
Google AI Studiogoogle-ai-studio/flex | Healthy | 100% | 99.96% | 99.88% | 99.68% | 1.82s | 7.26s | 149 tps | 1.0M | $0.375 / $1.88 |
Googlegoogle-vertex/global | Healthy | 99.68% | 99.34% | 99.07% | 99.39% | 2.24s | 10.9s | 92 tps | 1.0M | $0.750 / $3.75 |
Googlegoogle-vertex/global/flex | No data | — | 98.46% | 99.51% | 81.18% | 13.0s | 152s | 32 tps | 1.0M | $0.375 / $1.88 |
Googlegoogle-vertex/global/priority | No data | — | 99.81% | 99.80% | 99.76% | 2.76s | 5.64s | 98 tps | 1.0M | $1.35 / $6.75 |
Google AI Studiogoogle-ai-studio/priority | No data | — | 100% | 99.86% | 99.60% | 1.66s | 5.08s | 88 tps | 1.0M | $1.35 / $6.75 |
- Google AI Studio100%Healthy24h 99.88%
- 7d
- 99.60%
- 30d
- 99.51%
- p50
- 2.38s
- p99
- 10.2s
- Throughput
- 135 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-ai-studio - Google AI Studio100%Healthy24h 99.96%
- 7d
- 99.88%
- 30d
- 99.68%
- p50
- 1.82s
- p99
- 7.26s
- Throughput
- 149 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-ai-studio/flex - Google99.68%Healthy24h 99.34%
- 7d
- 99.07%
- 30d
- 99.39%
- p50
- 2.24s
- p99
- 10.9s
- Throughput
- 92 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-vertex/global - No data24h 98.46%
- 7d
- 99.51%
- 30d
- 81.18%
- p50
- 13.0s
- p99
- 152s
- Throughput
- 32 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-vertex/global/flex - No data24h 99.81%
- 7d
- 99.80%
- 30d
- 99.76%
- p50
- 2.76s
- p99
- 5.64s
- Throughput
- 98 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-vertex/global/priority - No data24h 100%
- 7d
- 99.86%
- 30d
- 99.60%
- p50
- 1.66s
- p99
- 5.08s
- Throughput
- 88 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-ai-studio/priority
Measured 3m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.
Throughput
Median tokens per second.
Daily availability
One row per endpoint · one dot per day · last 90 days.
worst hour ≥ 95%90–95%< 90%no data
Colour by each endpoint's worst hour of each UTC day · 6 endpoints · 47 days with data, 32 endpoint-days below 90%.
Price changes
- Gemini 3.7 FlashGooglegoogle-vertex/global$0.375 → $0.750 / $1.88 → $3.75cache read $0.037 → $0.075cache write $0.021 → $0.042reasoning $1.88 → $3.752026-08-28 18:10 UTC
- Gemini 3.7 FlashGooglegoogle-vertex/global/flex$0.188 → $0.375 / $0.938 → $1.88cache read $0.019 → $0.037cache write $0.010 → $0.021reasoning $0.938 → $1.882026-08-28 18:10 UTC
- Gemini 3.7 FlashGooglegoogle-vertex/global/priority$0.675 → $1.35 / $3.38 → $6.75cache read $0.068 → $0.135cache write $0.037 → $0.075reasoning $3.38 → $6.752026-08-28 18:10 UTC