OpenAI: gpt-oss-120b
by OpenAIOpenAI: gpt-oss-120b is a large language model from OpenAI. It costs $0.037 per million input tokens and $0.170 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.037
- Output / 1M tokens
- $0.170
- Cached input / 1M
- -
- Context window
- 131K
Not supported
tokens
Who serves it cheapest
19 hosts serve gpt-oss-120b. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| AkashMLCheapestbf16 | $0.030 | $0.170 | 131K | - | 100.0% |
| CoreWeavefp4 | $0.030 | $0.170 | 131K | - | 99.9% |
| DekaLLMbf16 | $0.030 | $0.180 | 131K | - | 99.3% |
| DeepInfrabf16 | $0.037 | $0.170 | 131K | - | 98.5% |
| Novitafp4 | $0.050 | $0.250 | 131K | - | 97.2% |
| DigitalOcean | $0.055 | $0.385 | 128K | - | 100.0% |
| $0.090 | $0.360 | 131K | - | 98.1% | |
| Mancer 2fp8 | $0.055 | $0.500 | 131K | - | 99.5% |
| BaseTenfp4 | $0.100 | $0.500 | 128K | - | 100.0% |
| Amazon Bedrock | $0.150 | $0.600 | 131K | - | 100.0% |
| Nebiusfp4 | $0.150 | $0.600 | 131K | - | 99.0% |
| SiliconFlowfp8 | $0.150 | $0.600 | 131K | - | 95.9% |
| Phala | $0.150 | $0.600 | 131K | - | 98.0% |
| Together | $0.150 | $0.600 | 131K | - | 88.5% |
| Groq | $0.150 | $0.600 | 131K | - | 100.0% |
| Parasailfp4 | $0.100 | $0.750 | 131K | - | 100.0% |
| Mara | $0.150 | $0.750 | 131K | - | 79.6% |
| SambaNova | $0.140 | $0.950 | 131K | - | 99.2% |
| Cerebrasfp16 | $0.350 | $0.750 | 131K | - | 100.0% |
The spread between AkashML and Cerebras is 6.9× for identical weights. Quantization and context limits differ, so check both columns before switching.
Benchmarks
Independent scores published alongside the catalogue.
- Intelligence index
- 12.3
- Coding index
- 30.4
- Agentic index
- 6.2
About gpt-oss-120b
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
Specifications
| Model ID | openai/gpt-oss-120b |
|---|---|
| Provider | OpenAI |
| Context window | 131K tokens |
| Max output | 118K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | 2024-06-30 |
| Open weights | Yes - openai/gpt-oss-120b |
| Released | August 5, 2025 |
Cheaper alternatives
Models that cost less than gpt-oss-120b while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does OpenAI: gpt-oss-120b cost?
$0.037 per million input tokens and $0.170 per million output tokens.
What is the context window of OpenAI: gpt-oss-120b?
131K tokens, with up to 118K tokens of output per request.
Which provider serves OpenAI: gpt-oss-120b cheapest?
AkashML at $0.030 per million input tokens - 6.9× cheaper than Cerebras, the most expensive of the 19 hosts serving it.
Confirm against the source: OpenAI official pricing.