Skip to content
LLMs
Qwen logo

Qwen: Qwen3 VL 32B Instruct

by Qwen

Qwen: Qwen3 VL 32B Instruct is a large language model from Qwen. It costs $0.104 per million input tokens and $0.416 per million output tokens. Its context window is 131K tokens.

Tool callingStructured outputOpen weightsimage input
Input / 1M tokens
$0.104
Output / 1M tokens
$0.416
Cached input / 1M
-

Not supported

Context window
131K

tokens

Who serves it cheapest

1 hosts serve Qwen3 VL 32B Instruct. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3 VL 32B Instruct, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
AlibabaCheapest$0.104$0.416131K-100.0%

About Qwen3 VL 32B Instruct

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

Specifications

Qwen: Qwen3 VL 32B Instruct specifications
Model IDqwen/qwen3-vl-32b-instruct
ProviderQwen
Context window131K tokens
Max output33K tokens
Input modalitiestext, image
Output modalitiestext
Knowledge cutoff-
Open weightsYes - Qwen/Qwen3-VL-32B-Instruct
ReleasedOctober 23, 2025

Cheaper alternatives

Models that cost less than Qwen3 VL 32B Instruct while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3 VL 32B Instruct cost?

$0.104 per million input tokens and $0.416 per million output tokens.

What is the context window of Qwen: Qwen3 VL 32B Instruct?

131K tokens, with up to 33K tokens of output per request.

Confirm against the source: Qwen official pricing.