OpenAI: GPT-4o-mini (2024-07-18)
by OpenAIOpenAI: GPT-4o-mini (2024-07-18) is a large language model from OpenAI. It costs $0.150 per million input tokens and $0.600 per million output tokens. Its context window is 128K tokens.
- Input / 1M tokens
- $0.150
- Output / 1M tokens
- $0.600
- Cached input / 1M
- $0.075
- Context window
- 128K
On repeated prefixes
tokens
Who serves it cheapest
1 hosts serve GPT-4o-mini (2024-07-18). Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| OpenAICheapest | $0.150 | $0.600 | 128K | - | 100.0% |
About GPT-4o-mini (2024-07-18)
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
Specifications
| Model ID | openai/gpt-4o-mini-2024-07-18 |
|---|---|
| Provider | OpenAI |
| Context window | 128K tokens |
| Max output | 16K tokens |
| Input modalities | text, image, file |
| Output modalities | text |
| Knowledge cutoff | 2023-10-31 |
| Open weights | No |
| Released | July 18, 2024 |
Cheaper alternatives
Models that cost less than GPT-4o-mini (2024-07-18) while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does OpenAI: GPT-4o-mini (2024-07-18) cost?
$0.150 per million input tokens and $0.600 per million output tokens. Cached input reads cost $0.075 per million tokens.
What is the context window of OpenAI: GPT-4o-mini (2024-07-18)?
128K tokens, with up to 16K tokens of output per request.
Confirm against the source: OpenAI official pricing.