mistralai logo
mistralai

Mistral: Voxtral Small 24B 2507

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. $0.10 per million input tokens, $0.30 per million output tokens. 32,768 token context window, maximum output of 32,768 tokens.

Overview

Model specifications

Context
32,768 tokens
Maximum output
26,214 tokens
Architecture
text+file+audio->text
Tokenizer
Mistral
Knowledge cutoff
Not provided
Moderated
No
OPENROUTER

Complete pricing

Synchronized OpenRouter rates. Token prices are shown per one million tokens.

Input
$0.1
/M tokens
Output
$0.3
/M tokens
Cache read
$0.01
/M tokens
Input Audio
$0.006
/minute
API

Model configuration

Default parameters

temperature
0.2
top_p
0.95
API

Quick start

Call this exact model ID through OpenRouter’s OpenAI-compatible API.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"mistralai/voxtral-small-24b-2507","messages":[{"role":"user","content":"Hello!"}]}'
Capabilities

InputOutput

Input
textaudiofile
Output
text
API

Supported API parameters

frequency_penaltymax_tokenspresence_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
Available providers

3 providers

Live providers on OpenRouter

Provider availability, latency, throughput and routing can change continuously. Open the source page for current operational data.

OpenRouter

Mistral

unknown
Context
33K
Maximum output
26K
Input
$0.100
Output
$0.300
Cache read
$0.010
Cache write
Not provided

Mistral

unknown
Context
32K
Maximum output
26K
Input
$0.100
Output
$0.300
Cache read
$0.010
Cache write
Not provided

Mistral

unknown
Context
32K
Maximum output
26K
Input
$0.110
Output
$0.330
Cache read
$0.011
Cache write
Not provided
mistralai

models

All models