Google: Gemini 2.5 Flash Lite
by GoogleGoogle: Gemini 2.5 Flash Lite is a large language model from Google. It costs $0.100 per million input tokens and $0.400 per million output tokens. Its context window is 1.0M tokens.
- Input / 1M tokens
- $0.100
- Output / 1M tokens
- $0.400
- Cached input / 1M
- $0.010
- Context window
- 1.0M
On repeated prefixes
tokens
This model also charges $0.014 per search request
That is $14.00 per 1,000 requests, billed on top of the token prices above. On a typical 1,000-in / 500-out exchange the tokens cost $0.0003, so the search fee is 47× the token cost. Per-token rankings elsewhere ignore this entirely.
Who serves it cheapest
2 hosts serve Gemini 2.5 Flash Lite. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| Google AI StudioCheapest | $0.050 | $0.200 | 1.0M | - | 100.0% |
| $0.100 | $0.400 | 1.0M | - | 99.9% |
The spread between Google AI Studio and Google is 2.0× for identical weights. Quantization and context limits differ, so check both columns before switching.
About Gemini 2.5 Flash Lite
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Specifications
| Model ID | google/gemini-2.5-flash-lite |
|---|---|
| Provider | |
| Context window | 1.0M tokens |
| Max output | 66K tokens |
| Input modalities | text, image, file, audio, video |
| Output modalities | text |
| Knowledge cutoff | 2025-01-31 |
| Open weights | No |
| Released | July 22, 2025 |
Cheaper alternatives
Models that cost less than Gemini 2.5 Flash Lite while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Google: Gemini 2.5 Flash Lite cost?
$0.100 per million input tokens and $0.400 per million output tokens. Cached input reads cost $0.010 per million tokens.
What is the context window of Google: Gemini 2.5 Flash Lite?
1.0M tokens, with up to 66K tokens of output per request.
Does Google: Gemini 2.5 Flash Lite have extra fees beyond per-token pricing?
Yes. Google: Gemini 2.5 Flash Lite charges $0.014 per search request - $14.00 per 1,000 requests - on top of $0.100 per million input tokens and $0.400 per million output tokens. On a typical 1,000-input / 500-output exchange the search fee alone is 47× the token cost.
Which provider serves Google: Gemini 2.5 Flash Lite cheapest?
Google AI Studio at $0.050 per million input tokens - 2.0× cheaper than Google, the most expensive of the 2 hosts serving it.
Confirm against the source: Google official pricing.