Google: Gemini 2.5 Flash
by GoogleGoogle: Gemini 2.5 Flash is a large language model from Google. It costs $0.300 per million input tokens and $2.50 per million output tokens. Its context window is 1.0M tokens.
- Input / 1M tokens
- $0.300
- Output / 1M tokens
- $2.50
- Cached input / 1M
- $0.030
- Context window
- 1.0M
On repeated prefixes
tokens
This model also charges $0.014 per search request
That is $14.00 per 1,000 requests, billed on top of the token prices above. On a typical 1,000-in / 500-out exchange the tokens cost $0.0015, so the search fee is 9.0× the token cost. Per-token rankings elsewhere ignore this entirely.
Who serves it cheapest
2 hosts serve Gemini 2.5 Flash. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| Google AI StudioCheapest | $0.150 | $1.25 | 1.0M | - | 100.0% |
| $0.300 | $2.50 | 1.0M | - | 99.0% |
The spread between Google AI Studio and Google is 2.0× for identical weights. Quantization and context limits differ, so check both columns before switching.
About Gemini 2.5 Flash
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Specifications
| Model ID | google/gemini-2.5-flash |
|---|---|
| Provider | |
| Context window | 1.0M tokens |
| Max output | 66K tokens |
| Input modalities | file, image, text, audio, video |
| Output modalities | text |
| Knowledge cutoff | 2025-01-31 |
| Open weights | No |
| Released | June 17, 2025 |
Cheaper alternatives
Models that cost less than Gemini 2.5 Flash while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Google: Gemini 2.5 Flash cost?
$0.300 per million input tokens and $2.50 per million output tokens. Cached input reads cost $0.030 per million tokens.
What is the context window of Google: Gemini 2.5 Flash?
1.0M tokens, with up to 66K tokens of output per request.
Does Google: Gemini 2.5 Flash have extra fees beyond per-token pricing?
Yes. Google: Gemini 2.5 Flash charges $0.014 per search request - $14.00 per 1,000 requests - on top of $0.300 per million input tokens and $2.50 per million output tokens. On a typical 1,000-input / 500-output exchange the search fee alone is 9.0× the token cost.
Which provider serves Google: Gemini 2.5 Flash cheapest?
Google AI Studio at $0.150 per million input tokens - 2.0× cheaper than Google, the most expensive of the 2 hosts serving it.
Confirm against the source: Google official pricing.