Skip to content
LLMs
Qwen logo

Qwen: Qwen3 VL 8B Thinking

by Qwen

Qwen: Qwen3 VL 8B Thinking is a large language model from Qwen. It costs $0.180 per million input tokens and $2.10 per million output tokens. Its context window is 131K tokens.

ReasoningTool callingStructured outputOpen weightsimage input
Input / 1M tokens
$0.180
Output / 1M tokens
$2.10
Cached input / 1M
-

Not supported

Context window
131K

tokens

Who serves it cheapest

1 hosts serve Qwen3 VL 8B Thinking. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3 VL 8B Thinking, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
AlibabaCheapest$0.180$2.10131K-100.0%

About Qwen3 VL 8B Thinking

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...

Specifications

Qwen: Qwen3 VL 8B Thinking specifications
Model IDqwen/qwen3-vl-8b-thinking
ProviderQwen
Context window131K tokens
Max output33K tokens
Input modalitiesimage, text
Output modalitiestext
Knowledge cutoff-
Open weightsYes - Qwen/Qwen3-VL-8B-Thinking
ReleasedOctober 14, 2025

Cheaper alternatives

Models that cost less than Qwen3 VL 8B Thinking while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3 VL 8B Thinking cost?

$0.180 per million input tokens and $2.10 per million output tokens.

What is the context window of Qwen: Qwen3 VL 8B Thinking?

131K tokens, with up to 33K tokens of output per request.

Confirm against the source: Qwen official pricing.