Qwen: Qwen3 VL 32B Instruct
by QwenQwen: Qwen3 VL 32B Instruct is a large language model from Qwen. It costs $0.104 per million input tokens and $0.416 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.104
- Output / 1M tokens
- $0.416
- Cached input / 1M
- -
- Context window
- 131K
Not supported
tokens
Who serves it cheapest
1 hosts serve Qwen3 VL 32B Instruct. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| AlibabaCheapest | $0.104 | $0.416 | 131K | - | 100.0% |
About Qwen3 VL 32B Instruct
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
Specifications
| Model ID | qwen/qwen3-vl-32b-instruct |
|---|---|
| Provider | Qwen |
| Context window | 131K tokens |
| Max output | 33K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | Yes - Qwen/Qwen3-VL-32B-Instruct |
| Released | October 23, 2025 |
Cheaper alternatives
Models that cost less than Qwen3 VL 32B Instruct while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Qwen: Qwen3 VL 32B Instruct cost?
$0.104 per million input tokens and $0.416 per million output tokens.
What is the context window of Qwen: Qwen3 VL 32B Instruct?
131K tokens, with up to 33K tokens of output per request.
Confirm against the source: Qwen official pricing.