Skip to content
LLMs
inclusionAI logo

inclusionAI: Ling 3.0 Flash Fin

by inclusionAI

inclusionAI: Ling 3.0 Flash Fin is a large language model from inclusionAI. It costs $0.060 per million input tokens and $0.180 per million output tokens. Its context window is 262K tokens.

Free tier availableReasoningTool callingStructured outputPrompt caching
Input / 1M tokens
$0.060
Output / 1M tokens
$0.180
Cached input / 1M
$0.012

On repeated prefixes

Context window
262K

tokens

Who serves it cheapest

1 hosts serve Ling 3.0 Flash Fin. Same weights, same API - the price difference is pure margin and routing.

Providers serving inclusionAI: Ling 3.0 Flash Fin, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
DeepInfraCheapestfp4$0.060$0.180262K-100.0%

About Ling 3.0 Flash Fin

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

Specifications

inclusionAI: Ling 3.0 Flash Fin specifications
Model IDinclusionai/ling-3.0-flash-fin
ProviderinclusionAI
Context window262K tokens
Max output236K tokens
Input modalitiestext
Output modalitiestext
Knowledge cutoff-
Open weightsNo
ReleasedAugust 27, 2026

Cheaper alternatives

Models that cost less than Ling 3.0 Flash Fin while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does inclusionAI: Ling 3.0 Flash Fin cost?

$0.060 per million input tokens and $0.180 per million output tokens. Cached input reads cost $0.012 per million tokens.

What is the context window of inclusionAI: Ling 3.0 Flash Fin?

262K tokens, with up to 236K tokens of output per request.