Qwen: Qwen3 VL 235B A22B Thinking
by QwenQwen: Qwen3 VL 235B A22B Thinking is a large language model from Qwen. It costs $0.400 per million input tokens and $4.00 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.400
- Output / 1M tokens
- $4.00
- Cached input / 1M
- -
- Context window
- 131K
Not supported
tokens
Who serves it cheapest
2 hosts serve Qwen3 VL 235B A22B Thinking. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| AlibabaCheapest | $0.400 | $4.00 | 131K | - | 91.4% |
| Novitabf16 | $0.980 | $3.95 | 131K | - | 99.6% |
The spread between Alibaba and Novita is 1.3× for identical weights. Quantization and context limits differ, so check both columns before switching.
About Qwen3 VL 235B A22B Thinking
Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....
Specifications
| Model ID | qwen/qwen3-vl-235b-a22b-thinking |
|---|---|
| Provider | Qwen |
| Context window | 131K tokens |
| Max output | 33K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Knowledge cutoff | 2025-03-31 |
| Open weights | Yes - Qwen/Qwen3-VL-235B-A22B-Thinking |
| Released | September 23, 2025 |
Cheaper alternatives
Models that cost less than Qwen3 VL 235B A22B Thinking while keeping at least half its context window and every input modality it supports.
Qwen3.5 397B A17B
$0.550 in · $3.50 out
about the same cheaperGrok Build 0.1
$1.00 in · $2.00 out
about the same cheaperSeed-2.0-Code
$0.500 in · $3.00 out
1.2× cheaperNano Banana 2 (Gemini 3.1 Flash Image)
$0.500 in · $3.00 out
1.2× cheaperNano Banana 2 (Gemini 3.1 Flash Image Preview)
$0.500 in · $3.00 out
1.2× cheaperFrequently asked
How much does Qwen: Qwen3 VL 235B A22B Thinking cost?
$0.400 per million input tokens and $4.00 per million output tokens.
What is the context window of Qwen: Qwen3 VL 235B A22B Thinking?
131K tokens, with up to 33K tokens of output per request.
Which provider serves Qwen: Qwen3 VL 235B A22B Thinking cheapest?
Alibaba at $0.400 per million input tokens - 1.3× cheaper than Novita, the most expensive of the 2 hosts serving it.
Confirm against the source: Qwen official pricing.