Qwen: Qwen3 Max
by QwenQwen: Qwen3 Max is a large language model from Qwen. It costs $0.780 per million input tokens and $3.90 per million output tokens. Its context window is 262K tokens.
- Input / 1M tokens
- $0.780
- Output / 1M tokens
- $3.90
- Cached input / 1M
- $0.156
- Context window
- 262K
On repeated prefixes
tokens
Who serves it cheapest
1 hosts serve Qwen3 Max. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| AlibabaCheapest | $0.780 | $3.90 | 262K | - | 100.0% |
About Qwen3 Max
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...
Specifications
| Model ID | qwen/qwen3-max |
|---|---|
| Provider | Qwen |
| Context window | 262K tokens |
| Max output | 66K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | 2025-06-30 |
| Open weights | No |
| Released | September 23, 2025 |
Cheaper alternatives
Models that cost less than Qwen3 Max while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Qwen: Qwen3 Max cost?
$0.780 per million input tokens and $3.90 per million output tokens. Cached input reads cost $0.156 per million tokens.
What is the context window of Qwen: Qwen3 Max?
262K tokens, with up to 66K tokens of output per request.
Confirm against the source: Qwen official pricing.