DeepSeek: DeepSeek V4 Flash Vision Exp
by DeepSeekDeepSeek: DeepSeek V4 Flash Vision Exp is a large language model from DeepSeek. It costs $0.220 per million input tokens and $0.660 per million output tokens. Its context window is 1.0M tokens.
- Input / 1M tokens
- $0.220
- Output / 1M tokens
- $0.660
- Cached input / 1M
- $0.0070
- Context window
- 1.0M
On repeated prefixes
tokens
Who serves it cheapest
6 hosts serve DeepSeek V4 Flash Vision Exp. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfraCheapestfp8 | $0.216 | $0.647 | 1.0M | - | 98.7% |
| Fireworks | $0.220 | $0.660 | 1.0M | - | 99.5% |
| GMICloudfp8 | $0.440 | $1.32 | 1.0M | - | 99.6% |
| SiliconFlowfp8 | $0.440 | $1.32 | 1.0M | - | 94.9% |
| AtlasCloudfp8 | $0.440 | $1.32 | 1.0M | - | 79.7% |
| Novita | $0.440 | $1.32 | 1.0M | - | 99.9% |
The spread between DeepInfra and Novita is 2.0× for identical weights. Quantization and context limits differ, so check both columns before switching.
About DeepSeek V4 Flash Vision Exp
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...
Specifications
| Model ID | deepseek/deepseek-v4-flash-vision-exp |
|---|---|
| Provider | DeepSeek |
| Context window | 1.0M tokens |
| Max output | 944K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | Yes - deepseek-ai/DeepSeek-V4-Flash-Vision-Exp |
| Released | August 21, 2026 |
Cheaper alternatives
Models that cost less than DeepSeek V4 Flash Vision Exp while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does DeepSeek: DeepSeek V4 Flash Vision Exp cost?
$0.220 per million input tokens and $0.660 per million output tokens. Cached input reads cost $0.0070 per million tokens.
What is the context window of DeepSeek: DeepSeek V4 Flash Vision Exp?
1.0M tokens, with up to 944K tokens of output per request.
Which provider serves DeepSeek: DeepSeek V4 Flash Vision Exp cheapest?
DeepInfra at $0.216 per million input tokens - 2.0× cheaper than Novita, the most expensive of the 6 hosts serving it.
Confirm against the source: DeepSeek official pricing.