Gemini 2.5 Flash
google/gemini-2.5-flash
Compare with Gemini 2.5 Flash Lite, Gemini 2.5 Flash (batch), Gemini 2.5 Pro
CompareGemini 2.5 Flash
| State | ||||||||||
|---|---|---|---|---|---|---|---|---|---|---|
Google AI Studiogoogle-ai-studio | Healthy | 99.94% | 99.75% | 99.89% | 99.89% | 597ms | 7.97s | 76 tps | 1.0M | $0.300 / $2.50 |
Google AI Studiogoogle-ai-studio/flex | Healthy | 99.94% | 99.75% | 99.89% | 99.89% | 2.36s | 5.36s | 4.0 tps | 1.0M | $0.150 / $1.25 |
Google AI Studiogoogle-ai-studio/priority | Healthy | 99.94% | 99.75% | 99.89% | 99.89% | — | — | — | 1.0M | $0.540 / $4.50 |
Googlegoogle-vertex/global | Healthy | 99.82% | 99.45% | 99.47% | 99.47% | 667ms | 5.56s | 73 tps | 1.0M | $0.300 / $2.50 |
Googlegoogle-vertex/global/priority | Healthy | 99.82% | 99.45% | 99.47% | 99.47% | 1.25s | 1.78s | 42 tps | 1.0M | $0.540 / $4.50 |
Googlegoogle-vertex/eu | Healthy | 99.14% | 98.28% | 87.57% | 87.57% | 717ms | 7.69s | 74 tps | 1.0M | $0.300 / $2.50 |
Googlegoogle-vertex | Down | 84.21% | 94.34% | 84.00% | 84.00% | 2.90s | 16.6s | 82 tps | 1.0M | $0.300 / $2.50 |
- Google AI Studio99.94%Healthy24h 99.75%
- 7d
- 99.89%
- 30d
- 99.89%
- p50
- 597ms
- p99
- 7.97s
- Throughput
- 76 tps
- Context
- 1.0M
- $/Mtok
- $0.300 / $2.50
google-ai-studio - Google AI Studio99.94%Healthy24h 99.75%
- 7d
- 99.89%
- 30d
- 99.89%
- p50
- 2.36s
- p99
- 5.36s
- Throughput
- 4.0 tps
- Context
- 1.0M
- $/Mtok
- $0.150 / $1.25
google-ai-studio/flex - Google AI Studio99.94%Healthy24h 99.75%
- 7d
- 99.89%
- 30d
- 99.89%
- p50
- —
- p99
- —
- Throughput
- —
- Context
- 1.0M
- $/Mtok
- $0.540 / $4.50
google-ai-studio/priority - Google99.82%Healthy24h 99.45%
- 7d
- 99.47%
- 30d
- 99.47%
- p50
- 667ms
- p99
- 5.56s
- Throughput
- 73 tps
- Context
- 1.0M
- $/Mtok
- $0.300 / $2.50
google-vertex/global - Google99.82%Healthy24h 99.45%
- 7d
- 99.47%
- 30d
- 99.47%
- p50
- 1.25s
- p99
- 1.78s
- Throughput
- 42 tps
- Context
- 1.0M
- $/Mtok
- $0.540 / $4.50
google-vertex/global/priority - Google99.14%Healthy24h 98.28%
- 7d
- 87.57%
- 30d
- 87.57%
- p50
- 717ms
- p99
- 7.69s
- Throughput
- 74 tps
- Context
- 1.0M
- $/Mtok
- $0.300 / $2.50
google-vertex/eu - Google84.21%Down24h 94.34%
- 7d
- 84.00%
- 30d
- 84.00%
- p50
- 2.90s
- p99
- 16.6s
- Throughput
- 82 tps
- Context
- 1.0M
- $/Mtok
- $0.300 / $2.50
google-vertex
Measured 3m ago. Latency and throughput are OpenRouter's rolling 30-minute windows.
History
Availability
Axis starts at 90% so small differences stay visible; it expands to 0 when an endpoint falls below.
Time to first token
Logarithmic axis — the fleet spans two orders of magnitude. P50 is OpenRouter's rolling 30-minute window.
Throughput
Median tokens per second.