Skip to content
LLMs
Qwen logo

Qwen: Qwen3 Coder Flash

by Qwen

Qwen: Qwen3 Coder Flash is a large language model from Qwen. It costs $0.195 per million input tokens and $0.975 per million output tokens. Its context window is 1M tokens.

Tool callingStructured outputPrompt caching
Input / 1M tokens
$0.195
Output / 1M tokens
$0.975
Cached input / 1M
$0.039

On repeated prefixes

Context window
1M

tokens

Who serves it cheapest

1 hosts serve Qwen3 Coder Flash. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3 Coder Flash, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
AlibabaCheapest$0.195$0.9751M-100.0%

About Qwen3 Coder Flash

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...

Specifications

Qwen: Qwen3 Coder Flash specifications
Model IDqwen/qwen3-coder-flash
ProviderQwen
Context window1M tokens
Max output66K tokens
Input modalitiestext
Output modalitiestext
Knowledge cutoff2025-06-30
Open weightsNo
ReleasedSeptember 17, 2025

Cheaper alternatives

Models that cost less than Qwen3 Coder Flash while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3 Coder Flash cost?

$0.195 per million input tokens and $0.975 per million output tokens. Cached input reads cost $0.039 per million tokens.

What is the context window of Qwen: Qwen3 Coder Flash?

1M tokens, with up to 66K tokens of output per request.

Confirm against the source: Qwen official pricing.