Skip to content
LLMs

Chinese AI models and what they cost

inclusionAI: Ling 3.0 Flash is the cheapest option at $0.021 per million input tokens. DeepSeek, Qwen, Kimi, GLM, MiniMax and the rest of the Chinese open-weight lineage. These models undercut Western equivalents by an order of magnitude and are the fastest-growing segment of the market, yet most English-language directories barely cover them.

121 models qualify

Ranked by blended cost - input weighted 75%, output 25%.

Language models with price per million tokens and context window
#ModelBlended / 1MInput / 1MOutput / 1MContextCapabilities
1FreeFreeFree262K
Free tierReasoningTools
2
inclusionAI logo
Ling 3.0 Flash

inclusionAI

$0.032$0.021$0.063262K
ReasoningToolsOpen weights
3$0.055$0.030$0.1301M
ReasoningToolsVision
4$0.075$0.060$0.1201.3M
ReasoningToolsOpen weights
5$0.077$0.044$0.1778K
Open weights
6$0.084$0.048$0.193262K
ToolsOpen weights
7$0.090$0.060$0.180131K
Free tierReasoningToolsVisionOpen weights
8$0.090$0.060$0.180262K
Free tierReasoningTools
9$0.107$0.086$0.1711.0M
ReasoningToolsOpen weights
10$0.112$0.100$0.150262K
ReasoningToolsVisionOpen weights
11$0.114$0.065$0.2601M
ReasoningToolsVision
12$0.123$0.070$0.280262K
ToolsOpen weights
13
ByteDance logo
UI-TARS 7B

ByteDance

$0.125$0.100$0.200128K
VisionOpen weights
14$0.125$0.100$0.20033K
ToolsOpen weights
15$0.129$0.074$0.2958K
Open weights
16
Tencent logo
Hy-MT2-7B

Tencent

$0.129$0.074$0.2958K
Open weights
17$0.130$0.080$0.280131K
ReasoningToolsOpen weights
18
ByteDance Seed logo
Seed 1.6 Flash

ByteDance Seed

$0.131$0.075$0.300262K
ReasoningToolsVision
19$0.145$0.061$0.400200K
ReasoningToolsOpen weights
20$0.150$0.100$0.300262K
ReasoningToolsOpen weights
21$0.150$0.120$0.240131K
ReasoningToolsOpen weights
22$0.153$0.087$0.350262K
ToolsOpen weights
23
ByteDance Seed logo
Seed-2.0-Mini

ByteDance Seed

$0.175$0.100$0.400262K
ReasoningToolsVision
24
Xiaomi logo
MiMo-V2.5

Xiaomi

$0.175$0.140$0.2801.1M
ReasoningToolsVisionOpen weights
25$0.182$0.104$0.416131K
ToolsVisionOpen weights

Frequently asked

What is the cheapest LLM for Chinese-developed models?

inclusionAI: Ling 3.0 Flash at $0.021 per million input tokens and $0.063 per million output tokens, with a 262K token context window.

How many models qualify for Chinese-developed models?

121 of the 340 models tracked here meet the criteria for this list.

How much do prices vary within this category?

By roughly 170×, from Ling 3.0 Flash at the bottom to Kimi K3 at the top.