Skip to content
LLMs
Mistral AI logo

Mistral: Voxtral Small 24B 2507

by Mistral AI

Mistral: Voxtral Small 24B 2507 is a large language model from Mistral AI. It costs $0.100 per million input tokens and $0.300 per million output tokens. Its context window is 33K tokens.

Tool callingStructured outputPrompt cachingOpen weightsaudio inputfile input
Input / 1M tokens
$0.100
Output / 1M tokens
$0.300
Cached input / 1M
$0.010

On repeated prefixes

Context window
33K

tokens

Who serves it cheapest

1 hosts serve Voxtral Small 24B 2507. Same weights, same API - the price difference is pure margin and routing.

Providers serving Mistral: Voxtral Small 24B 2507, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
MistralCheapest$0.100$0.30033K-99.6%

About Voxtral Small 24B 2507

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...

Specifications

Mistral: Voxtral Small 24B 2507 specifications
Model IDmistralai/voxtral-small-24b-2507
ProviderMistral AI
Context window33K tokens
Max output26K tokens
Input modalitiestext, audio, file
Output modalitiestext
Knowledge cutoff-
Open weightsYes - mistralai/Voxtral-Small-24B-2507
ReleasedOctober 30, 2025

Cheaper alternatives

Models that cost less than Voxtral Small 24B 2507 while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Mistral: Voxtral Small 24B 2507 cost?

$0.100 per million input tokens and $0.300 per million output tokens. Cached input reads cost $0.010 per million tokens.

What is the context window of Mistral: Voxtral Small 24B 2507?

33K tokens, with up to 26K tokens of output per request.

Confirm against the source: Mistral AI official pricing.