Skip to content
LLMs
inclusionAI logo

inclusionAI: Ling 3.0 Flash VL

by inclusionAI

inclusionAI: Ling 3.0 Flash VL is a large language model from inclusionAI. It costs $0.060 per million input tokens and $0.180 per million output tokens. Its context window is 131K tokens.

Free tier availableReasoningTool callingStructured outputPrompt cachingOpen weightsimage inputvideo input
Input / 1M tokens
$0.060
Output / 1M tokens
$0.180
Cached input / 1M
$0.012

On repeated prefixes

Context window
131K

tokens

Who serves it cheapest

1 hosts serve Ling 3.0 Flash VL. Same weights, same API - the price difference is pure margin and routing.

Providers serving inclusionAI: Ling 3.0 Flash VL, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
DeepInfraCheapestfp16$0.060$0.180131K-100.0%

Benchmarks

Independent scores published alongside the catalogue.

Intelligence index
24.8
Coding index
57.0
Agentic index
30.0

About Ling 3.0 Flash VL

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

Specifications

inclusionAI: Ling 3.0 Flash VL specifications
Model IDinclusionai/ling-3.0-flash-vl
ProviderinclusionAI
Context window131K tokens
Max output33K tokens
Input modalitiestext, image, video
Output modalitiestext
Knowledge cutoff-
Open weightsYes - inclusionAI/Ling-3.0-flash-VL
ReleasedSeptember 10, 2026

Cheaper alternatives

Models that cost less than Ling 3.0 Flash VL while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does inclusionAI: Ling 3.0 Flash VL cost?

$0.060 per million input tokens and $0.180 per million output tokens. Cached input reads cost $0.012 per million tokens.

What is the context window of inclusionAI: Ling 3.0 Flash VL?

131K tokens, with up to 33K tokens of output per request.