Qwen: Qwen3.7 Flash
by QwenQwen: Qwen3.7 Flash is a large language model from Qwen. It costs $0.030 per million input tokens and $0.130 per million output tokens. Its context window is 1M tokens.
- Input / 1M tokens
- $0.030
- Output / 1M tokens
- $0.130
- Cached input / 1M
- $0.0060
- Context window
- 1M
On repeated prefixes
tokens
Who serves it cheapest
1 hosts serve Qwen3.7 Flash. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| AlibabaCheapest | $0.030 | $0.130 | 1M | - | 100.0% |
About Qwen3.7 Flash
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
Specifications
| Model ID | qwen/qwen3.7-flash |
|---|---|
| Provider | Qwen |
| Context window | 1M tokens |
| Max output | 66K tokens |
| Input modalities | text, image, video |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | No |
| Released | July 27, 2026 |
Cheaper alternatives
Models that cost less than Qwen3.7 Flash while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Qwen: Qwen3.7 Flash cost?
$0.030 per million input tokens and $0.130 per million output tokens. Cached input reads cost $0.0060 per million tokens.
What is the context window of Qwen: Qwen3.7 Flash?
1M tokens, with up to 66K tokens of output per request.
Confirm against the source: Qwen official pricing.