Qwen: Qwen3 30B A3B
by QwenQwen: Qwen3 30B A3B is a large language model from Qwen. It costs $0.120 per million input tokens and $0.500 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.120
- Output / 1M tokens
- $0.500
- Cached input / 1M
- -
- Context window
- 131K
Not supported
tokens
Who serves it cheapest
2 hosts serve Qwen3 30B A3B. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfraCheapestfp8 | $0.120 | $0.500 | 41K | - | 99.9% |
| Alibaba | $0.130 | $0.520 | 131K | - | 100.0% |
The spread between DeepInfra and Alibaba is 1.1× for identical weights. Quantization and context limits differ, so check both columns before switching.
About Qwen3 30B A3B
Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...
Specifications
| Model ID | qwen/qwen3-30b-a3b |
|---|---|
| Provider | Qwen |
| Context window | 131K tokens |
| Max output | 16K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | 2025-03-31 |
| Open weights | Yes - Qwen/Qwen3-30B-A3B |
| Released | April 28, 2025 |
Cheaper alternatives
Models that cost less than Qwen3 30B A3B while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Qwen: Qwen3 30B A3B cost?
$0.120 per million input tokens and $0.500 per million output tokens.
What is the context window of Qwen: Qwen3 30B A3B?
131K tokens, with up to 16K tokens of output per request.
Which provider serves Qwen: Qwen3 30B A3B cheapest?
DeepInfra at $0.120 per million input tokens - 1.1× cheaper than Alibaba, the most expensive of the 2 hosts serving it.
Confirm against the source: Qwen official pricing.