mistralai logo
mistralai

Mistral: Voxtral Mini TTS

Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. Priced at $16 per M characters. 4,096 token context window. Includes independent benchmarks from Artificial Analysis.

模型概览

模型规格

上下文
4,096 tokens
最大输出
3,276 tokens
架构
text->speech
分词器
Mistral
知识截止时间
暂未提供
内容审核
OPENROUTER

完整价格

同步自 OpenRouter。Token 价格均按每百万 Token 展示。

Characters
$16
/M characters
API

模型配置

支持的声音

en_paul_saden_paul_neutralen_paul_happyen_paul_frustrateden_paul_exciteden_paul_confidenten_paul_cheerfulen_paul_angrygb_oliver_neutralgb_oliver_sadgb_oliver_excitedgb_oliver_curiousgb_oliver_confidentgb_oliver_cheerfulgb_oliver_angrygb_jane_sarcasmgb_jane_confusedgb_jane_shamefulgb_jane_sadgb_jane_neutralgb_jane_jealousygb_jane_frustratedgb_jane_curiousgb_jane_confidentfr_marie_sadfr_marie_neutralfr_marie_happyfr_marie_excitedfr_marie_curiousfr_marie_angry
API

快速调用

通过 OpenRouter 的 OpenAI 兼容 API 调用这个确切的模型 ID。

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"mistralai/voxtral-mini-tts-2603","messages":[{"role":"user","content":"Hello!"}]}'
能力与模态

输入输出

输入
text
输出
speech
API

支持的 API 参数

frequency_penaltymax_tokenspresence_penaltyresponse_formatseedstopstructured_outputstemperaturetop_p
可用提供方

3 个提供方

在 OpenRouter 查看实时 Provider

Provider 可用性、延迟、吞吐量和路由会持续变化,请前往来源页面查看实时运行数据。

OpenRouter

Mistral

unknown
上下文
4K
最大输出
3K
输入
$16.00
输出
免费
缓存读取
暂未提供
缓存写入
暂未提供

Mistral

unknown
上下文
4K
最大输出
3K
输入
$16.00
输出
免费
缓存读取
暂未提供
缓存写入
暂未提供

Mistral

unknown
上下文
4K
最大输出
3K
输入
$17.60
输出
免费
缓存读取
暂未提供
缓存写入
暂未提供
mistralai

个模型

全部模型
mistralai logomistralai

Mistral: Voxtral Small 24B 2507 STT

Voxtral Small 24B 2507 STT is a speech transcription model from Mistral AI. It is suited for transcription, translation, and audio understanding workloads that benefit from its larger model capacity.

上下文
暂未提供
输入
audio
输出
transcription
输入: $50.00输出: 免费每百万 Token
查看模型详情
mistralai logomistralai

Mistral: Voxtral Mini 3B 2507

Voxtral Mini 3B 2507 is a speech and audio understanding model from Mistral AI. It is suited for transcription, translation, and compact audio processing workloads.

上下文
暂未提供
输入
audio
输出
transcription
输入: $16.67输出: 免费每百万 Token
查看模型详情
mistralai logomistralai

Mistral: Voxtral Mini Transcribe

Voxtral Mini Transcribe is Mistral's speech-to-text model, derived from the Voxtral Mini family. It accepts audio input and returns transcribed text via the standard transcription API. Suited for transcribing meetings,...

上下文
暂未提供
输入
audio
输出
transcription
输入: $3000.00输出: 免费每百万 Token
查看模型详情
mistralai logomistralai

Mistral: Mistral Medium 3.5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

上下文
262K
输入
text · image · file
输出
text
输入: $1.50输出: $7.50每百万 Token
查看模型详情