Qwen: Qwen3 8B
by QwenQwen: Qwen3 8B is a large language model from Qwen. It costs $0.117 per million input tokens and $0.455 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.117
- Output / 1M tokens
- $0.455
- Cached input / 1M
- -
- Context window
- 131K
Not supported
tokens
Who serves it cheapest
1 hosts serve Qwen3 8B. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| AlibabaCheapest | $0.117 | $0.455 | 131K | - | 100.0% |
Benchmarks
Independent scores published alongside the catalogue.
- Intelligence index
- 5.2
- Coding index
- 9.0
- Agentic index
- 0.8
About Qwen3 8B
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
Specifications
| Model ID | qwen/qwen3-8b |
|---|---|
| Provider | Qwen |
| Context window | 131K tokens |
| Max output | 8K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | 2025-03-31 |
| Open weights | Yes - Qwen/Qwen3-8B |
| Released | April 28, 2025 |
Cheaper alternatives
Models that cost less than Qwen3 8B while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Qwen: Qwen3 8B cost?
$0.117 per million input tokens and $0.455 per million output tokens.
What is the context window of Qwen: Qwen3 8B?
131K tokens, with up to 8K tokens of output per request.
Confirm against the source: Qwen official pricing.