fish-audio logo
fish-audio

Fish Audio: S2 Pro

Beschreibung der Quelle (Englisch)

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

Überblick

Modellspezifikationen

Kontext
Nicht angegeben
Maximale Ausgabe
Nicht angegeben
Architektur
text->speech
Tokenizer
Other
Wissensstand
Nicht angegeben
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

UTF-8-Bytes
$15
pro Million UTF-8-Bytes
API

Schnellstart

Setzen Sie OPENROUTER_API_KEY lokal. Python benötigt requests; JavaScript läuft in Node.js. Bewahren Sie den Schlüssel auf dem Server auf.

Wählen Sie eine unterstützte Stimme. Ersetzen Sie YOUR_VOICE_ID, falls keine Stimme gelistet ist.

API-Dokumentation

curl --fail-with-body https://openrouter.ai/api/v1/audio/speech \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"fish-audio/s2-pro","input":"Hello!","voice":"YOUR_VOICE_ID","response_format":"mp3"}' \
  --output speech.mp3
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
text
Ausgabe
speech
API

Unterstützte API-Parameter

Verfügbare Anbieter

1 Anbieter

Geprüft: 7. September 2026

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

Fish Audio

Nicht angegeben
Kontext
Nicht angegeben
Maximale Ausgabe
Nicht angegeben
fish-audio

4 Modelle

Alle Modelle
fish-audio logofish-audio

Fish Audio: Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

Kontext
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Audiodauer: $0.0001 pro Sekunde
Modelldetails ansehen
fish-audio logofish-audio

Fish Audio: S1

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

Kontext
Nicht angegeben
Eingabe
text
Ausgabe
speech
UTF-8-Bytes: $15 pro Million UTF-8-Bytes
Modelldetails ansehen
fish-audio logofish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

Kontext
Nicht angegeben
Eingabe
text
Ausgabe
speech
Eingabe: Kostenlos pro 1 Mio. TokenAusgabe: Kostenlos pro 1 Mio. Token
Modelldetails ansehen
fish-audio logofish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

Kontext
Nicht angegeben
Eingabe
text
Ausgabe
speech
UTF-8-Bytes: $15 pro Million UTF-8-Bytes
Modelldetails ansehen