Mistral: Codestral 2508
by Mistral AIMistral: Codestral 2508 is a large language model from Mistral AI. It costs $0.300 per million input tokens and $0.900 per million output tokens. Its context window is 256K tokens.
- Input / 1M tokens
- $0.300
- Output / 1M tokens
- $0.900
- Cached input / 1M
- $0.030
- Context window
- 256K
On repeated prefixes
tokens
Who serves it cheapest
1 hosts serve Codestral 2508. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| MistralCheapest | $0.300 | $0.900 | 256K | - | 99.9% |
About Codestral 2508
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
Specifications
| Model ID | mistralai/codestral-2508 |
|---|---|
| Provider | Mistral AI |
| Context window | 256K tokens |
| Max output | 205K tokens |
| Input modalities | text, file |
| Output modalities | text |
| Knowledge cutoff | 2025-03-31 |
| Open weights | No |
| Released | August 1, 2025 |
Cheaper alternatives
Models that cost less than Codestral 2508 while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Mistral: Codestral 2508 cost?
$0.300 per million input tokens and $0.900 per million output tokens. Cached input reads cost $0.030 per million tokens.
What is the context window of Mistral: Codestral 2508?
256K tokens, with up to 205K tokens of output per request.
Confirm against the source: Mistral AI official pricing.