Qwen: Qwen3 VL 235B A22B Instruct
by QwenQwen: Qwen3 VL 235B A22B Instruct is a large language model from Qwen. It costs $0.210 per million input tokens and $1.90 per million output tokens. Its context window is 262K tokens.
- Input / 1M tokens
- $0.210
- Output / 1M tokens
- $1.90
- Cached input / 1M
- $0.100
- Context window
- 262K
On repeated prefixes
tokens
Who serves it cheapest
5 hosts serve Qwen3 VL 235B A22B Instruct. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfraCheapestfp8 | $0.200 | $0.880 | 262K | - | 94.1% |
| Alibaba | $0.260 | $1.04 | 131K | - | 99.6% |
| Novitabf16 | $0.300 | $1.50 | 131K | - | 97.0% |
| Venicefp8 | $0.210 | $1.90 | 128K | - | 95.7% |
| Parasailfp8 | $0.210 | $1.90 | 131K | - | 99.4% |
The spread between DeepInfra and Parasail is 1.7× for identical weights. Quantization and context limits differ, so check both columns before switching.
About Qwen3 VL 235B A22B Instruct
Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table...
Specifications
| Model ID | qwen/qwen3-vl-235b-a22b-instruct |
|---|---|
| Provider | Qwen |
| Context window | 262K tokens |
| Max output | 33K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Knowledge cutoff | 2025-03-31 |
| Open weights | Yes - Qwen/Qwen3-VL-235B-A22B-Instruct |
| Released | September 23, 2025 |
Cheaper alternatives
Models that cost less than Qwen3 VL 235B A22B Instruct while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Qwen: Qwen3 VL 235B A22B Instruct cost?
$0.210 per million input tokens and $1.90 per million output tokens. Cached input reads cost $0.100 per million tokens.
What is the context window of Qwen: Qwen3 VL 235B A22B Instruct?
262K tokens, with up to 33K tokens of output per request.
Which provider serves Qwen: Qwen3 VL 235B A22B Instruct cheapest?
DeepInfra at $0.200 per million input tokens - 1.7× cheaper than Parasail, the most expensive of the 5 hosts serving it.
Confirm against the source: Qwen official pricing.