Skip to content
LLMs
Qwen logo

Qwen: Qwen3 8B

by Qwen

Qwen: Qwen3 8B is a large language model from Qwen. It costs $0.117 per million input tokens and $0.455 per million output tokens. Its context window is 131K tokens.

ReasoningTool callingStructured outputOpen weights
Input / 1M tokens
$0.117
Output / 1M tokens
$0.455
Cached input / 1M
-

Not supported

Context window
131K

tokens

Who serves it cheapest

1 hosts serve Qwen3 8B. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3 8B, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
AlibabaCheapest$0.117$0.455131K-100.0%

Benchmarks

Independent scores published alongside the catalogue.

Intelligence index
5.2
Coding index
9.0
Agentic index
0.8

About Qwen3 8B

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

Specifications

Qwen: Qwen3 8B specifications
Model IDqwen/qwen3-8b
ProviderQwen
Context window131K tokens
Max output8K tokens
Input modalitiestext
Output modalitiestext
Knowledge cutoff2025-03-31
Open weightsYes - Qwen/Qwen3-8B
ReleasedApril 28, 2025

Cheaper alternatives

Models that cost less than Qwen3 8B while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3 8B cost?

$0.117 per million input tokens and $0.455 per million output tokens.

What is the context window of Qwen: Qwen3 8B?

131K tokens, with up to 8K tokens of output per request.

Confirm against the source: Qwen official pricing.