Qwen: Qwen3 30B A3B Instruct 2507
by QwenQwen: Qwen3 30B A3B Instruct 2507 is a large language model from Qwen. It costs $0.048 per million input tokens and $0.193 per million output tokens. Its context window is 262K tokens.
- Input / 1M tokens
- $0.048
- Output / 1M tokens
- $0.193
- Cached input / 1M
- -
- Context window
- 262K
Not supported
tokens
Who serves it cheapest
5 hosts serve Qwen3 30B A3B Instruct 2507. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| StreamLakeCheapest | $0.048 | $0.193 | 128K | - | 97.9% |
| DekaLLM | $0.090 | $0.300 | 262K | - | 99.3% |
| SiliconFlowfp8 | $0.090 | $0.300 | 262K | - | 97.5% |
| Nebiusfp8 | $0.100 | $0.300 | 262K | - | 92.9% |
| Alibaba | $0.130 | $0.520 | 131K | - | 100.0% |
The spread between StreamLake and Alibaba is 2.7× for identical weights. Quantization and context limits differ, so check both columns before switching.
About Qwen3 30B A3B Instruct 2507
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...
Specifications
| Model ID | qwen/qwen3-30b-a3b-instruct-2507 |
|---|---|
| Provider | Qwen |
| Context window | 262K tokens |
| Max output | 32K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | 2025-06-30 |
| Open weights | Yes - Qwen/Qwen3-30B-A3B-Instruct-2507 |
| Released | July 29, 2025 |
Cheaper alternatives
Models that cost less than Qwen3 30B A3B Instruct 2507 while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Qwen: Qwen3 30B A3B Instruct 2507 cost?
$0.048 per million input tokens and $0.193 per million output tokens.
What is the context window of Qwen: Qwen3 30B A3B Instruct 2507?
262K tokens, with up to 32K tokens of output per request.
Which provider serves Qwen: Qwen3 30B A3B Instruct 2507 cheapest?
StreamLake at $0.048 per million input tokens - 2.7× cheaper than Alibaba, the most expensive of the 5 hosts serving it.
Confirm against the source: Qwen official pricing.