inclusionAI: Ling 3.0 Flash Fin
by inclusionAIinclusionAI: Ling 3.0 Flash Fin is a large language model from inclusionAI. It costs $0.060 per million input tokens and $0.180 per million output tokens. Its context window is 262K tokens.
- Input / 1M tokens
- $0.060
- Output / 1M tokens
- $0.180
- Cached input / 1M
- $0.012
- Context window
- 262K
On repeated prefixes
tokens
Who serves it cheapest
1 hosts serve Ling 3.0 Flash Fin. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfraCheapestfp4 | $0.060 | $0.180 | 262K | - | 100.0% |
About Ling 3.0 Flash Fin
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
Specifications
| Model ID | inclusionai/ling-3.0-flash-fin |
|---|---|
| Provider | inclusionAI |
| Context window | 262K tokens |
| Max output | 236K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | No |
| Released | August 27, 2026 |
Cheaper alternatives
Models that cost less than Ling 3.0 Flash Fin while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does inclusionAI: Ling 3.0 Flash Fin cost?
$0.060 per million input tokens and $0.180 per million output tokens. Cached input reads cost $0.012 per million tokens.
What is the context window of inclusionAI: Ling 3.0 Flash Fin?
262K tokens, with up to 236K tokens of output per request.