Skip to content
LLMs
Qwen logo

Qwen: Qwen3.5-9B

by Qwen

Qwen: Qwen3.5-9B is a large language model from Qwen. It costs $0.100 per million input tokens and $0.150 per million output tokens. Its context window is 262K tokens.

ReasoningTool callingStructured outputBatch tierOpen weightsimage inputvideo input
Input / 1M tokens
$0.100
Output / 1M tokens
$0.150
Cached input / 1M
-

Not supported

Context window
262K

tokens

Who serves it cheapest

6 hosts serve Qwen3.5-9B. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3.5-9B, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
DarkbloomCheapestfp4$0.080$0.130262K-99.9%
SiliconFlowfp8$0.100$0.150262K-97.4%
DeepInfrabf16$0.100$0.150262K-98.0%
Venicefp8$0.100$0.150256K-99.5%
Parasailbf16$0.100$0.250262K-86.9%
Together$0.170$0.250262K-99.4%

The spread between Darkbloom and Together is 2.1× for identical weights. Quantization and context limits differ, so check both columns before switching.

Benchmarks

Independent scores published alongside the catalogue.

Coding index
28.7

About Qwen3.5-9B

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

Specifications

Qwen: Qwen3.5-9B specifications
Model IDqwen/qwen3.5-9b
ProviderQwen
Context window262K tokens
Max output236K tokens
Input modalitiestext, image, video
Output modalitiestext
Knowledge cutoff-
Open weightsYes - Qwen/Qwen3.5-9B
ReleasedMarch 10, 2026

Cheaper alternatives

Models that cost less than Qwen3.5-9B while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3.5-9B cost?

$0.100 per million input tokens and $0.150 per million output tokens.

What is the context window of Qwen: Qwen3.5-9B?

262K tokens, with up to 236K tokens of output per request.

Which provider serves Qwen: Qwen3.5-9B cheapest?

Darkbloom at $0.080 per million input tokens - 2.1× cheaper than Together, the most expensive of the 6 hosts serving it.

Confirm against the source: Qwen official pricing.