Qwen: Qwen3 235B A22B Instruct 2507
by QwenQwen: Qwen3 235B A22B Instruct 2507 is a large language model from Qwen. It costs $0.087 per million input tokens and $0.350 per million output tokens. Its context window is 262K tokens.
- Input / 1M tokens
- $0.087
- Output / 1M tokens
- $0.350
- Cached input / 1M
- $0.018
- Context window
- 262K
On repeated prefixes
tokens
Who serves it cheapest
10 hosts serve Qwen3 235B A22B Instruct 2507. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| GMICloudCheapestfp8 | $0.087 | $0.350 | 262K | - | 97.2% |
| DeepInfrafp8 | $0.090 | $0.550 | 262K | - | 96.6% |
| Novitafp8 | $0.090 | $0.580 | 131K | - | 98.5% |
| Alibaba | $0.150 | $0.598 | 131K | - | 100.0% |
| Venicefp8 | $0.150 | $0.750 | 128K | - | 97.0% |
| Nebiusfp8 | $0.200 | $0.600 | 262K | - | 75.8% |
| Parasailfp8 | $0.140 | $0.800 | 131K | - | 99.9% |
| StreamLake | $0.210 | $0.840 | 128K | - | 96.5% |
| AtlasCloudfp8 | $0.200 | $0.880 | 131K | - | 96.2% |
| $0.220 | $0.880 | 262K | - | 99.9% |
The spread between GMICloud and Google is 2.5× for identical weights. Quantization and context limits differ, so check both columns before switching.
About Qwen3 235B A22B Instruct 2507
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
Specifications
| Model ID | qwen/qwen3-235b-a22b-2507 |
|---|---|
| Provider | Qwen |
| Context window | 262K tokens |
| Max output | 236K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | 2025-06-30 |
| Open weights | Yes - Qwen/Qwen3-235B-A22B-Instruct-2507 |
| Released | July 21, 2025 |
Cheaper alternatives
Models that cost less than Qwen3 235B A22B Instruct 2507 while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Qwen: Qwen3 235B A22B Instruct 2507 cost?
$0.087 per million input tokens and $0.350 per million output tokens. Cached input reads cost $0.018 per million tokens.
What is the context window of Qwen: Qwen3 235B A22B Instruct 2507?
262K tokens, with up to 236K tokens of output per request.
Which provider serves Qwen: Qwen3 235B A22B Instruct 2507 cheapest?
GMICloud at $0.087 per million input tokens - 2.5× cheaper than Google, the most expensive of the 10 hosts serving it.
Confirm against the source: Qwen official pricing.