Chinese AI models and what they cost
inclusionAI: Ling 3.0 Flash is the cheapest option at $0.021 per million input tokens. DeepSeek, Qwen, Kimi, GLM, MiniMax and the rest of the Chinese open-weight lineage. These models undercut Western equivalents by an order of magnitude and are the fastest-growing segment of the market, yet most English-language directories barely cover them.
121 models qualify
Ranked by blended cost - input weighted 75%, output 25%.
| # | Model | Blended / 1M | Input / 1M | Output / 1M | Context | Capabilities |
|---|---|---|---|---|---|---|
| 1 | Ling 3.0 Flash Sante (free) inclusionAI | Free | Free | Free | 262K | Free tierReasoningTools |
| 2 | Ling 3.0 Flash inclusionAI | $0.032 | $0.021 | $0.063 | 262K | ReasoningToolsOpen weights |
| 3 | Qwen3.7 Flash Qwen | $0.055 | $0.030 | $0.130 | 1M | ReasoningToolsVision |
| 4 | DeepSeek V4 Flash 0731 DeepSeek | $0.075 | $0.060 | $0.120 | 1.3M | ReasoningToolsOpen weights |
| 5 | Hy-MT2-1.8B Tencent | $0.077 | $0.044 | $0.177 | 8K | Open weights |
| 6 | $0.084 | $0.048 | $0.193 | 262K | ToolsOpen weights | |
| 7 | Ling 3.0 Flash VL inclusionAI | $0.090 | $0.060 | $0.180 | 131K | Free tierReasoningToolsVisionOpen weights |
| 8 | Ling 3.0 Flash Fin inclusionAI | $0.090 | $0.060 | $0.180 | 262K | Free tierReasoningTools |
| 9 | DeepSeek V4 Flash 0423 DeepSeek | $0.107 | $0.086 | $0.171 | 1.0M | ReasoningToolsOpen weights |
| 10 | Qwen3.5-9B Qwen | $0.112 | $0.100 | $0.150 | 262K | ReasoningToolsVisionOpen weights |
| 11 | Qwen3.5-Flash Qwen | $0.114 | $0.065 | $0.260 | 1M | ReasoningToolsVision |
| 12 | $0.123 | $0.070 | $0.280 | 262K | ToolsOpen weights | |
| 13 | UI-TARS 7B ByteDance | $0.125 | $0.100 | $0.200 | 128K | VisionOpen weights |
| 14 | $0.125 | $0.100 | $0.200 | 33K | ToolsOpen weights | |
| 15 | Hy-MT2-30B-A3B Tencent | $0.129 | $0.074 | $0.295 | 8K | Open weights |
| 16 | Hy-MT2-7B Tencent | $0.129 | $0.074 | $0.295 | 8K | Open weights |
| 17 | Qwen3 32B Qwen | $0.130 | $0.080 | $0.280 | 131K | ReasoningToolsOpen weights |
| 18 | Seed 1.6 Flash ByteDance Seed | $0.131 | $0.075 | $0.300 | 262K | ReasoningToolsVision |
| 19 | GLM 4.7 Flash Z.ai | $0.145 | $0.061 | $0.400 | 200K | ReasoningToolsOpen weights |
| 20 | Step 3.5 Flash StepFun | $0.150 | $0.100 | $0.300 | 262K | ReasoningToolsOpen weights |
| 21 | Qwen3 14B Qwen | $0.150 | $0.120 | $0.240 | 131K | ReasoningToolsOpen weights |
| 22 | $0.153 | $0.087 | $0.350 | 262K | ToolsOpen weights | |
| 23 | Seed-2.0-Mini ByteDance Seed | $0.175 | $0.100 | $0.400 | 262K | ReasoningToolsVision |
| 24 | MiMo-V2.5 Xiaomi | $0.175 | $0.140 | $0.280 | 1.1M | ReasoningToolsVisionOpen weights |
| 25 | $0.182 | $0.104 | $0.416 | 131K | ToolsVisionOpen weights |
Frequently asked
What is the cheapest LLM for Chinese-developed models?
inclusionAI: Ling 3.0 Flash at $0.021 per million input tokens and $0.063 per million output tokens, with a 262K token context window.
How many models qualify for Chinese-developed models?
121 of the 340 models tracked here meet the criteria for this list.
How much do prices vary within this category?
By roughly 170×, from Ling 3.0 Flash at the bottom to Kimi K3 at the top.