mistralai logo
mistralai

Mistral: Voxtral Mini TTS

Opis źródłowy (angielski)

Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. It converts text input into natural-sounding audio output.

Przegląd

Specyfikacja modelu

Kontekst
4096 tokens
Maksymalne wyjście
3276 tokens
Architektura
text->speech
Tokenizer
Mistral
Granica wiedzy
Brak danych
Moderowany
Nie
OPENROUTER

Pełny cennik

Stawki zsynchronizowane z OpenRouter, za milion tokenów.

Znaki
$16
za milion znaków
API

Konfiguracja modelu

Obsługiwane głosy

en_paul_saden_paul_neutralen_paul_happyen_paul_frustrateden_paul_exciteden_paul_confidenten_paul_cheerfulen_paul_angrygb_oliver_neutralgb_oliver_sadgb_oliver_excitedgb_oliver_curiousgb_oliver_confidentgb_oliver_cheerfulgb_oliver_angrygb_jane_sarcasmgb_jane_confusedgb_jane_shamefulgb_jane_sadgb_jane_neutralgb_jane_jealousygb_jane_frustratedgb_jane_curiousgb_jane_confidentfr_marie_sadfr_marie_neutralfr_marie_happyfr_marie_excitedfr_marie_curiousfr_marie_angry
API

Szybki start

Ustaw lokalnie OPENROUTER_API_KEY. Python wymaga requests, a JavaScript działa w Node.js. Klucz przechowuj na serwerze.

Wybierz obsługiwany głos. Zastąp YOUR_VOICE_ID, jeśli nie podano głosu.

Dokumentacja API

curl --fail-with-body https://openrouter.ai/api/v1/audio/speech \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"mistralai/voxtral-mini-tts-2603","input":"Hello!","voice":"en_paul_sad","response_format":"mp3"}' \
  --output speech.mp3
Możliwości i modalności

WejścieWyjście

Wejście
text
Wyjście
speech
API

Obsługiwane parametry API

Dostępni dostawcy

3 dostawców

Sprawdzono: 16 września 2026

Dostawcy na żywo w OpenRouter

Dostępność, opóźnienie, przepustowość i routing stale się zmieniają. Aktualne dane są na stronie źródłowej.

OpenRouter

Mistral

Brak danych
Kontekst
4K
Maksymalne wyjście
3K

Mistral

Brak danych
Kontekst
4K
Maksymalne wyjście
3K

Mistral

Brak danych
Kontekst
4K
Maksymalne wyjście
3K
mistralai

4 modele

Wszystkie modele
mistralai logomistralai

Mistral: Voxtral Small 24B 2507 STT

Voxtral Small 24B 2507 STT is a speech transcription model from Mistral AI. It is suited for transcription, translation, and audio understanding workloads that benefit from its larger model capacity.

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Czas audio: $0.00005 za sekundę
Zobacz szczegóły modelu
mistralai logomistralai

Mistral: Voxtral Mini 3B 2507

Voxtral Mini 3B 2507 is a speech and audio understanding model from Mistral AI. It is suited for transcription, translation, and compact audio processing workloads.

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Czas audio: $0.000017 za sekundę
Zobacz szczegóły modelu
mistralai logomistralai

Mistral: Voxtral Mini Transcribe

Voxtral Mini Transcribe is Mistral's speech-to-text model, derived from the Voxtral Mini family. It accepts audio input and returns transcribed text via the standard transcription API. Suited for transcribing meetings,...

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Czas audio: $0.00005 za sekundę
Zobacz szczegóły modelu
mistralai logomistralai

Mistral: Mistral Medium 3.5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Kontekst
262K
Wejście
text · image · file
Wyjście
text
Wejście: $1.5Wyjście: $7.5za milion tokenów
Zobacz szczegóły modelu