Qwen: Qwen2.5 VL 72B Instruct
by QwenQwen: Qwen2.5 VL 72B Instruct is a large language model from Qwen. It costs $0.800 per million input tokens and $1.00 per million output tokens. Its context window is 128K tokens.
- Input / 1M tokens
- $0.800
- Output / 1M tokens
- $1.00
- Cached input / 1M
- $0.400
- Context window
- 128K
On repeated prefixes
tokens
Who serves it cheapest
1 hosts serve Qwen2.5 VL 72B Instruct. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| ParasailCheapestfp8 | $0.800 | $1.00 | 128K | - | 100.0% |
About Qwen2.5 VL 72B Instruct
Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.
Specifications
| Model ID | qwen/qwen2.5-vl-72b-instruct |
|---|---|
| Provider | Qwen |
| Context window | 128K tokens |
| Max output | 115K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Knowledge cutoff | 2024-06-30 |
| Open weights | Yes - Qwen/Qwen2.5-VL-72B-Instruct |
| Released | February 1, 2025 |
Cheaper alternatives
Models that cost less than Qwen2.5 VL 72B Instruct while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Qwen: Qwen2.5 VL 72B Instruct cost?
$0.800 per million input tokens and $1.00 per million output tokens. Cached input reads cost $0.400 per million tokens.
What is the context window of Qwen: Qwen2.5 VL 72B Instruct?
128K tokens, with up to 115K tokens of output per request.
Confirm against the source: Qwen official pricing.