Skip to content
LLMs
Qwen logo

Qwen: Qwen3.5-Flash

by Qwen

Qwen: Qwen3.5-Flash is a large language model from Qwen. It costs $0.065 per million input tokens and $0.260 per million output tokens. Its context window is 1M tokens.

ReasoningTool callingStructured outputimage inputvideo input
Input / 1M tokens
$0.065
Output / 1M tokens
$0.260
Cached input / 1M
-

Not supported

Context window
1M

tokens

Who serves it cheapest

1 hosts serve Qwen3.5-Flash. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3.5-Flash, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
AlibabaCheapest$0.065$0.2601M-100.0%

About Qwen3.5-Flash

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

Specifications

Qwen: Qwen3.5-Flash specifications
Model IDqwen/qwen3.5-flash-02-23
ProviderQwen
Context window1M tokens
Max output66K tokens
Input modalitiestext, image, video
Output modalitiestext
Knowledge cutoff-
Open weightsNo
ReleasedFebruary 25, 2026

Cheaper alternatives

Models that cost less than Qwen3.5-Flash while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3.5-Flash cost?

$0.065 per million input tokens and $0.260 per million output tokens.

What is the context window of Qwen: Qwen3.5-Flash?

1M tokens, with up to 66K tokens of output per request.

Confirm against the source: Qwen official pricing.