Skip to content
LLMs
OpenAI logo

OpenAI: gpt-oss-20b

by OpenAI

OpenAI: gpt-oss-20b is a large language model from OpenAI. It costs $0.030 per million input tokens and $0.130 per million output tokens. Its context window is 131K tokens.

ReasoningTool callingStructured outputPrompt cachingBatch tierOpen weights
Input / 1M tokens
$0.030
Output / 1M tokens
$0.130
Cached input / 1M
$0.030

On repeated prefixes

Context window
131K

tokens

Who serves it cheapest

13 hosts serve gpt-oss-20b. Same weights, same API - the price difference is pure margin and routing.

Providers serving OpenAI: gpt-oss-20b, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
DarkbloomCheapestfp8$0.020$0.100131K-99.2%
AkashMLfp4$0.020$0.100131K-99.7%
CoreWeavefp4$0.030$0.130131K-100.0%
DekaLLMbf16$0.029$0.140131K-99.7%
DeepInfrabf16$0.030$0.140131K-100.0%
Parasailfp4$0.030$0.150131K-99.9%
Phala$0.040$0.150131K-99.6%
Novitafp4$0.040$0.150131K-99.9%
SiliconFlowfp8$0.040$0.180131K-98.1%
Together$0.050$0.200131K-94.3%
Amazon Bedrock$0.070$0.150131K-99.9%
Google$0.070$0.250131K-99.9%
Groq$0.075$0.300131K-99.8%

The spread between Darkbloom and Groq is 3.3× for identical weights. Quantization and context limits differ, so check both columns before switching.

Benchmarks

Independent scores published alongside the catalogue.

Intelligence index
9.0
Coding index
20.7
Agentic index
1.4

About gpt-oss-20b

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

Specifications

OpenAI: gpt-oss-20b specifications
Model IDopenai/gpt-oss-20b
ProviderOpenAI
Context window131K tokens
Max output118K tokens
Input modalitiestext
Output modalitiestext
Knowledge cutoff2024-06-30
Open weightsYes - openai/gpt-oss-20b
ReleasedAugust 5, 2025

Cheaper alternatives

Models that cost less than gpt-oss-20b while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does OpenAI: gpt-oss-20b cost?

$0.030 per million input tokens and $0.130 per million output tokens. Cached input reads cost $0.030 per million tokens.

What is the context window of OpenAI: gpt-oss-20b?

131K tokens, with up to 118K tokens of output per request.

Which provider serves OpenAI: gpt-oss-20b cheapest?

Darkbloom at $0.020 per million input tokens - 3.3× cheaper than Groq, the most expensive of the 13 hosts serving it.

Confirm against the source: OpenAI official pricing.