Gemini 3.7 Flash
google/gemini-3.7-flashInputOutput
Compare its 6 providersIncident RSSCompare with Gemini 3.8 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash
CompareGemini 3.7 Flash
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
Google AI Studiogoogle-ai-studio/flex | Healthy | 100% | 99.95% | 99.88% | 99.69% | 1.88s | 7.97s | 152 tps | 1.0M | $0.375 / $1.88 |
Google AI Studiogoogle-ai-studio | Healthy | 99.17% | 99.89% | 99.60% | 99.52% | 2.15s | 9.01s | 129 tps | 1.0M | $0.750 / $3.75 |
Googlegoogle-vertex/global | Healthy | 95.00% | 99.30% | 99.05% | 99.39% | 2.53s | 15.8s | 88 tps | 1.0M | $0.750 / $3.75 |
Googlegoogle-vertex/global/flex | No data | — | 98.41% | 99.50% | 81.34% | 16.7s | 227s | 42 tps | 1.0M | $0.375 / $1.88 |
Googlegoogle-vertex/global/priority | No data | — | 99.87% | 99.80% | 99.76% | 2.14s | 5.05s | 85 tps | 1.0M | $1.35 / $6.75 |
Google AI Studiogoogle-ai-studio/priority | No data | — | 100% | 99.86% | 99.60% | 1.89s | 7.57s | 80 tps | 1.0M | $1.35 / $6.75 |
- Google AI Studio100%Healthy24h 99.95%
- 7d
- 99.88%
- 30d
- 99.69%
- p50
- 1.88s
- p99
- 7.97s
- Throughput
- 152 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-ai-studio/flex - Google AI Studio99.17%Healthy24h 99.89%
- 7d
- 99.60%
- 30d
- 99.52%
- p50
- 2.15s
- p99
- 9.01s
- Throughput
- 129 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-ai-studio - Google95.00%Healthy24h 99.30%
- 7d
- 99.05%
- 30d
- 99.39%
- p50
- 2.53s
- p99
- 15.8s
- Throughput
- 88 tps
- Context
- 1.0M
- $/Mtok
- $0.750 / $3.75
google-vertex/global - No data24h 98.41%
- 7d
- 99.50%
- 30d
- 81.34%
- p50
- 16.7s
- p99
- 227s
- Throughput
- 42 tps
- Context
- 1.0M
- $/Mtok
- $0.375 / $1.88
google-vertex/global/flex - No data24h 99.87%
- 7d
- 99.80%
- 30d
- 99.76%
- p50
- 2.14s
- p99
- 5.05s
- Throughput
- 85 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-vertex/global/priority - No data24h 100%
- 7d
- 99.86%
- 30d
- 99.60%
- p50
- 1.89s
- p99
- 7.57s
- Throughput
- 80 tps
- Context
- 1.0M
- $/Mtok
- $1.35 / $6.75
google-ai-studio/priority
Measured 2m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window, averaged over the hour.
Throughput
Median tokens per second.
Daily availability
One row per endpoint · one dot per day · last 90 days.
worst hour ≥ 95%90–95%< 90%no data
Colour by each endpoint's worst hour of each UTC day · 6 endpoints · 47 days with data, 33 endpoint-days below 90%.
Price changes
- Gemini 3.7 FlashGooglegoogle-vertex/global$0.375 → $0.750 / $1.88 → $3.75cache read $0.037 → $0.075cache write $0.021 → $0.042reasoning $1.88 → $3.752026-08-28 18:10 UTC
- Gemini 3.7 FlashGooglegoogle-vertex/global/flex$0.188 → $0.375 / $0.938 → $1.88cache read $0.019 → $0.037cache write $0.010 → $0.021reasoning $0.938 → $1.882026-08-28 18:10 UTC
- Gemini 3.7 FlashGooglegoogle-vertex/global/priority$0.675 → $1.35 / $3.38 → $6.75cache read $0.068 → $0.135cache write $0.037 → $0.075reasoning $3.38 → $6.752026-08-28 18:10 UTC