Google: Gemini 3.5 Flash
by GoogleGoogle: Gemini 3.5 Flash is a large language model from Google. It costs $1.50 per million input tokens and $9.00 per million output tokens. Its context window is 1.0M tokens.
- Input / 1M tokens
- $1.50
- Output / 1M tokens
- $9.00
- Cached input / 1M
- $0.150
- Context window
- 1.0M
On repeated prefixes
tokens
This model also charges $0.014 per search request
That is $14.00 per 1,000 requests, billed on top of the token prices above. On a typical 1,000-in / 500-out exchange the tokens cost $0.006, so the search fee is 2.3× the token cost. Per-token rankings elsewhere ignore this entirely.
Who serves it cheapest
2 hosts serve Gemini 3.5 Flash. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| GoogleCheapest | $0.750 | $4.50 | 1.0M | - | 75.6% |
| Google AI Studio | $0.750 | $4.50 | 1.0M | - | 100.0% |
The spread between Google and Google AI Studio is about the same for identical weights. Quantization and context limits differ, so check both columns before switching.
Benchmarks
Independent scores published alongside the catalogue.
- Intelligence index
- 33.0
- Coding index
- 70.1
- Agentic index
- 27.3
About Gemini 3.5 Flash
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Specifications
| Model ID | google/gemini-3.5-flash |
|---|---|
| Provider | |
| Context window | 1.0M tokens |
| Max output | 66K tokens |
| Input modalities | text, image, video, file, audio |
| Output modalities | text |
| Knowledge cutoff | 2025-01-01 |
| Open weights | No |
| Released | May 19, 2026 |
Cheaper alternatives
Models that cost less than Gemini 3.5 Flash while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Google: Gemini 3.5 Flash cost?
$1.50 per million input tokens and $9.00 per million output tokens. Cached input reads cost $0.150 per million tokens.
What is the context window of Google: Gemini 3.5 Flash?
1.0M tokens, with up to 66K tokens of output per request.
Does Google: Gemini 3.5 Flash have extra fees beyond per-token pricing?
Yes. Google: Gemini 3.5 Flash charges $0.014 per search request - $14.00 per 1,000 requests - on top of $1.50 per million input tokens and $9.00 per million output tokens. On a typical 1,000-input / 500-output exchange the search fee alone is 2.3× the token cost.
Which provider serves Google: Gemini 3.5 Flash cheapest?
Google at $0.750 per million input tokens - about the same cheaper than Google AI Studio, the most expensive of the 2 hosts serving it.
Confirm against the source: Google official pricing.