fish-audio logo
fish-audio

Fish Audio: Transcribe 1

Beschreibung der Quelle (Englisch)

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

Überblick

Modellspezifikationen

Kontext
Nicht angegeben
Maximale Ausgabe
Nicht angegeben
Architektur
audio->transcription
Tokenizer
Other
Wissensstand
Nicht angegeben
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Audiodauer
$0.0001
pro Sekunde
API

Schnellstart

Setzen Sie OPENROUTER_API_KEY lokal. Python benötigt requests; JavaScript läuft in Node.js. Bewahren Sie den Schlüssel auf dem Server auf.

API-Dokumentation

curl --fail-with-body https://openrouter.ai/api/v1/audio/transcriptions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -F 'model=fish-audio/transcribe-1' \
  -F 'file=@audio.wav'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
audio
Ausgabe
transcription
API

Unterstützte API-Parameter

Verfügbare Anbieter

1 Anbieter

Geprüft: 7. September 2026

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

Fish Audio

Nicht angegeben
Kontext
Nicht angegeben
Maximale Ausgabe
Nicht angegeben
fish-audio

4 Modelle

Alle Modelle
fish-audio logofish-audio

Fish Audio: S1

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

Kontext
Nicht angegeben
Eingabe
text
Ausgabe
speech
UTF-8-Bytes: $15 pro Million UTF-8-Bytes
Modelldetails ansehen
fish-audio logofish-audio

Fish Audio: S2 Pro

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

Kontext
Nicht angegeben
Eingabe
text
Ausgabe
speech
UTF-8-Bytes: $15 pro Million UTF-8-Bytes
Modelldetails ansehen
fish-audio logofish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

Kontext
Nicht angegeben
Eingabe
text
Ausgabe
speech
Eingabe: Kostenlos pro 1 Mio. TokenAusgabe: Kostenlos pro 1 Mio. Token
Modelldetails ansehen
fish-audio logofish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

Kontext
Nicht angegeben
Eingabe
text
Ausgabe
speech
UTF-8-Bytes: $15 pro Million UTF-8-Bytes
Modelldetails ansehen