Skip to content
LLMs
Qwen logo

Qwen: Qwen3 Max

by Qwen

Qwen: Qwen3 Max is a large language model from Qwen. It costs $0.780 per million input tokens and $3.90 per million output tokens. Its context window is 262K tokens.

Tool callingStructured outputPrompt caching
Input / 1M tokens
$0.780
Output / 1M tokens
$3.90
Cached input / 1M
$0.156

On repeated prefixes

Context window
262K

tokens

Who serves it cheapest

1 hosts serve Qwen3 Max. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3 Max, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
AlibabaCheapest$0.780$3.90262K-100.0%

About Qwen3 Max

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...

Specifications

Qwen: Qwen3 Max specifications
Model IDqwen/qwen3-max
ProviderQwen
Context window262K tokens
Max output66K tokens
Input modalitiestext
Output modalitiestext
Knowledge cutoff2025-06-30
Open weightsNo
ReleasedSeptember 23, 2025

Cheaper alternatives

Models that cost less than Qwen3 Max while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3 Max cost?

$0.780 per million input tokens and $3.90 per million output tokens. Cached input reads cost $0.156 per million tokens.

What is the context window of Qwen: Qwen3 Max?

262K tokens, with up to 66K tokens of output per request.

Confirm against the source: Qwen official pricing.