Skip to content
LLMs
Qwen logo

Qwen: Qwen3.7 Flash

by Qwen

Qwen: Qwen3.7 Flash is a large language model from Qwen. It costs $0.030 per million input tokens and $0.130 per million output tokens. Its context window is 1M tokens.

ReasoningTool callingStructured outputPrompt cachingimage inputvideo input
Input / 1M tokens
$0.030
Output / 1M tokens
$0.130
Cached input / 1M
$0.0060

On repeated prefixes

Context window
1M

tokens

Who serves it cheapest

1 hosts serve Qwen3.7 Flash. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3.7 Flash, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
AlibabaCheapest$0.030$0.1301M-100.0%

About Qwen3.7 Flash

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

Specifications

Qwen: Qwen3.7 Flash specifications
Model IDqwen/qwen3.7-flash
ProviderQwen
Context window1M tokens
Max output66K tokens
Input modalitiestext, image, video
Output modalitiestext
Knowledge cutoff-
Open weightsNo
ReleasedJuly 27, 2026

Cheaper alternatives

Models that cost less than Qwen3.7 Flash while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3.7 Flash cost?

$0.030 per million input tokens and $0.130 per million output tokens. Cached input reads cost $0.0060 per million tokens.

What is the context window of Qwen: Qwen3.7 Flash?

1M tokens, with up to 66K tokens of output per request.

Confirm against the source: Qwen official pricing.