mistralai logo
mistralai

Mistral: Voxtral Mini 3B 2507

Beschreibung der Quelle (Englisch)

Voxtral Mini 3B 2507 is a speech and audio understanding model from Mistral AI. It is suited for transcription, translation, and compact audio processing workloads.

Überblick

Modellspezifikationen

Kontext
Nicht angegeben
Maximale Ausgabe
Nicht angegeben
Architektur
audio->transcription
Tokenizer
Mistral
Wissensstand
Nicht angegeben
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Audiodauer
$0.000017
pro Sekunde
API

Schnellstart

Setzen Sie OPENROUTER_API_KEY lokal. Python benötigt requests; JavaScript läuft in Node.js. Bewahren Sie den Schlüssel auf dem Server auf.

API-Dokumentation

curl --fail-with-body https://openrouter.ai/api/v1/audio/transcriptions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -F 'model=mistralai/voxtral-mini-3b-2507' \
  -F 'file=@audio.wav'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
audio
Ausgabe
transcription
API

Unterstützte API-Parameter

Verfügbare Anbieter

1 Anbieter

Geprüft: 16. September 2026

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

DeepInfra

bf16
Kontext
Nicht angegeben
Maximale Ausgabe
Nicht angegeben
mistralai

4 Modelle

Alle Modelle
mistralai logomistralai

Mistral: Voxtral Small 24B 2507 STT

Voxtral Small 24B 2507 STT is a speech transcription model from Mistral AI. It is suited for transcription, translation, and audio understanding workloads that benefit from its larger model capacity.

Kontext
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Audiodauer: $0.00005 pro Sekunde
Modelldetails ansehen
mistralai logomistralai

Mistral: Voxtral Mini Transcribe

Voxtral Mini Transcribe is Mistral's speech-to-text model, derived from the Voxtral Mini family. It accepts audio input and returns transcribed text via the standard transcription API. Suited for transcribing meetings,...

Kontext
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Audiodauer: $0.00005 pro Sekunde
Modelldetails ansehen
mistralai logomistralai

Mistral: Mistral Medium 3.5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Kontext
262K
Eingabe
text · image · file
Ausgabe
text
Eingabe: $1.5Ausgabe: $7.5pro 1 Mio. Token
Modelldetails ansehen
mistralai logomistralai

Mistral: Mistral Medium 3.5 (batch)

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Kontext
262K
Eingabe
text · image · file
Ausgabe
text
Eingabe: $0.75Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen