Skip to content
LLMs
Qwen logo

Qwen: Qwen3 235B A22B Instruct 2507

by Qwen

Qwen: Qwen3 235B A22B Instruct 2507 is a large language model from Qwen. It costs $0.087 per million input tokens and $0.350 per million output tokens. Its context window is 262K tokens.

Tool callingStructured outputPrompt cachingOpen weights
Input / 1M tokens
$0.087
Output / 1M tokens
$0.350
Cached input / 1M
$0.018

On repeated prefixes

Context window
262K

tokens

Who serves it cheapest

10 hosts serve Qwen3 235B A22B Instruct 2507. Same weights, same API - the price difference is pure margin and routing.

Providers serving Qwen: Qwen3 235B A22B Instruct 2507, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
GMICloudCheapestfp8$0.087$0.350262K-97.2%
DeepInfrafp8$0.090$0.550262K-96.6%
Novitafp8$0.090$0.580131K-98.5%
Alibaba$0.150$0.598131K-100.0%
Venicefp8$0.150$0.750128K-97.0%
Nebiusfp8$0.200$0.600262K-75.8%
Parasailfp8$0.140$0.800131K-99.9%
StreamLake$0.210$0.840128K-96.5%
AtlasCloudfp8$0.200$0.880131K-96.2%
Google$0.220$0.880262K-99.9%

The spread between GMICloud and Google is 2.5× for identical weights. Quantization and context limits differ, so check both columns before switching.

About Qwen3 235B A22B Instruct 2507

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

Specifications

Qwen: Qwen3 235B A22B Instruct 2507 specifications
Model IDqwen/qwen3-235b-a22b-2507
ProviderQwen
Context window262K tokens
Max output236K tokens
Input modalitiestext
Output modalitiestext
Knowledge cutoff2025-06-30
Open weightsYes - Qwen/Qwen3-235B-A22B-Instruct-2507
ReleasedJuly 21, 2025

Cheaper alternatives

Models that cost less than Qwen3 235B A22B Instruct 2507 while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Qwen: Qwen3 235B A22B Instruct 2507 cost?

$0.087 per million input tokens and $0.350 per million output tokens. Cached input reads cost $0.018 per million tokens.

What is the context window of Qwen: Qwen3 235B A22B Instruct 2507?

262K tokens, with up to 236K tokens of output per request.

Which provider serves Qwen: Qwen3 235B A22B Instruct 2507 cheapest?

GMICloud at $0.087 per million input tokens - 2.5× cheaper than Google, the most expensive of the 10 hosts serving it.

Confirm against the source: Qwen official pricing.