mistralai logo
mistralai

Mistral: Voxtral Mini TTS

Descrição da fonte (inglês)

Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. It converts text input into natural-sounding audio output.

Visão geral

Especificações do modelo

Contexto
4.096 tokens
Saída máxima
3.276 tokens
Arquitetura
text->speech
Tokenizador
Mistral
Limite de conhecimento
Não informado
Moderado
Não
OPENROUTER

Preços completos

Tarifas sincronizadas do OpenRouter, por milhão de tokens.

Caracteres
$16
por milhão de caracteres
API

Configuração do modelo

Vozes compatíveis

en_paul_saden_paul_neutralen_paul_happyen_paul_frustrateden_paul_exciteden_paul_confidenten_paul_cheerfulen_paul_angrygb_oliver_neutralgb_oliver_sadgb_oliver_excitedgb_oliver_curiousgb_oliver_confidentgb_oliver_cheerfulgb_oliver_angrygb_jane_sarcasmgb_jane_confusedgb_jane_shamefulgb_jane_sadgb_jane_neutralgb_jane_jealousygb_jane_frustratedgb_jane_curiousgb_jane_confidentfr_marie_sadfr_marie_neutralfr_marie_happyfr_marie_excitedfr_marie_curiousfr_marie_angry
API

Início rápido

Defina OPENROUTER_API_KEY localmente. Python requer requests; JavaScript roda em Node.js. Mantenha a chave no servidor.

Escolha uma voz compatível. Substitua YOUR_VOICE_ID se nenhuma voz estiver listada.

Documentação da API

curl --fail-with-body https://openrouter.ai/api/v1/audio/speech \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"mistralai/voxtral-mini-tts-2603","input":"Hello!","voice":"en_paul_sad","response_format":"mp3"}' \
  --output speech.mp3
Recursos e modalidades

EntradaSaída

Entrada
text
Saída
speech
API

Parâmetros de API compatíveis

Provedores disponíveis

3 provedores

Verificado: 16 de setembro de 2026

Provedores ao vivo no OpenRouter

Disponibilidade, latência, throughput e roteamento mudam continuamente. Consulte a fonte para dados atuais.

OpenRouter

Mistral

Não informado
Contexto
4K
Saída máxima
3K

Mistral

Não informado
Contexto
4K
Saída máxima
3K

Mistral

Não informado
Contexto
4K
Saída máxima
3K
mistralai

4 modelos

Todos os modelos
mistralai logomistralai

Mistral: Voxtral Small 24B 2507 STT

Voxtral Small 24B 2507 STT is a speech transcription model from Mistral AI. It is suited for transcription, translation, and audio understanding workloads that benefit from its larger model capacity.

Contexto
Não informado
Entrada
audio
Saída
transcription
Duração do áudio: $0.00005 por segundo
Ver detalhes do modelo
mistralai logomistralai

Mistral: Voxtral Mini 3B 2507

Voxtral Mini 3B 2507 is a speech and audio understanding model from Mistral AI. It is suited for transcription, translation, and compact audio processing workloads.

Contexto
Não informado
Entrada
audio
Saída
transcription
Duração do áudio: $0.000017 por segundo
Ver detalhes do modelo
mistralai logomistralai

Mistral: Voxtral Mini Transcribe

Voxtral Mini Transcribe is Mistral's speech-to-text model, derived from the Voxtral Mini family. It accepts audio input and returns transcribed text via the standard transcription API. Suited for transcribing meetings,...

Contexto
Não informado
Entrada
audio
Saída
transcription
Duração do áudio: $0.00005 por segundo
Ver detalhes do modelo
mistralai logomistralai

Mistral: Mistral Medium 3.5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Contexto
262K
Entrada
text · image · file
Saída
text
Entrada: $1.5Saída: $7.5por milhão de tokens
Ver detalhes do modelo