Skip to content
LLMs
Qwen logo

Qwen: Qwen3 VL 30B A3B Instruct

by Qwen

Qwen: Qwen3 VL 30B A3B Instruct is a large language model from Qwen. It costs $0.150 per million input tokens and $0.600 per million output tokens. Its context window is 262K tokens.

Tool callingStructured outputOpen weightsimage input
Input / 1M tokens
$0.150
Output / 1M tokens
$0.600
Cached input / 1M
-

Not supported

Context window
262K

tokens

Who serves it cheapest

4 hosts serve Qwen3 VL 30B A3B Instruct. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3 VL 30B A3B Instruct, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
AlibabaCheapest$0.130$0.520131K-100.0%
DeepInfrafp8$0.150$0.600262K-95.4%
Novitabf16$0.200$0.700131K-97.1%
SiliconFlowfp8$0.290$1.00262K-93.6%

The spread between Alibaba and SiliconFlow is 2.1× for identical weights. Quantization and context limits differ, so check both columns before switching.

About Qwen3 VL 30B A3B Instruct

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...

Specifications

Qwen: Qwen3 VL 30B A3B Instruct specifications
Model IDqwen/qwen3-vl-30b-a3b-instruct
ProviderQwen
Context window262K tokens
Max output16K tokens
Input modalitiestext, image
Output modalitiestext
Knowledge cutoff2025-03-31
Open weightsYes - Qwen/Qwen3-VL-30B-A3B-Instruct
ReleasedOctober 6, 2025

Cheaper alternatives

Models that cost less than Qwen3 VL 30B A3B Instruct while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3 VL 30B A3B Instruct cost?

$0.150 per million input tokens and $0.600 per million output tokens.

What is the context window of Qwen: Qwen3 VL 30B A3B Instruct?

262K tokens, with up to 16K tokens of output per request.

Which provider serves Qwen: Qwen3 VL 30B A3B Instruct cheapest?

Alibaba at $0.130 per million input tokens - 2.1× cheaper than SiliconFlow, the most expensive of the 4 hosts serving it.

Confirm against the source: Qwen official pricing.