inclusionAI: Ling 3.0 Flash VL
by inclusionAIinclusionAI: Ling 3.0 Flash VL is a large language model from inclusionAI. It costs $0.060 per million input tokens and $0.180 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.060
- Output / 1M tokens
- $0.180
- Cached input / 1M
- $0.012
- Context window
- 131K
On repeated prefixes
tokens
Who serves it cheapest
1 hosts serve Ling 3.0 Flash VL. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfraCheapestfp16 | $0.060 | $0.180 | 131K | - | 100.0% |
Benchmarks
Independent scores published alongside the catalogue.
- Intelligence index
- 24.8
- Coding index
- 57.0
- Agentic index
- 30.0
About Ling 3.0 Flash VL
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
Specifications
| Model ID | inclusionai/ling-3.0-flash-vl |
|---|---|
| Provider | inclusionAI |
| Context window | 131K tokens |
| Max output | 33K tokens |
| Input modalities | text, image, video |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | Yes - inclusionAI/Ling-3.0-flash-VL |
| Released | September 10, 2026 |
Cheaper alternatives
Models that cost less than Ling 3.0 Flash VL while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does inclusionAI: Ling 3.0 Flash VL cost?
$0.060 per million input tokens and $0.180 per million output tokens. Cached input reads cost $0.012 per million tokens.
What is the context window of inclusionAI: Ling 3.0 Flash VL?
131K tokens, with up to 33K tokens of output per request.