FA
fish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

Überblick

Modellspezifikationen

Context
Nicht angegeben
Maximale Ausgabe
Nicht angegeben
Architektur
text->speech
Tokenizer
Other
Wissensstand
Nicht angegeben
Moderiert
Nein
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
text
Ausgabe
speech
API

Unterstützte API-Parameter

Verfügbare Anbieter

1 Anbieter

Fish Audio

unknown
Context
Nicht angegeben
Maximale Ausgabe
Nicht angegeben
Eingabe
$15.00
Ausgabe
Kostenlos
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben
fish-audio

models

Alle Modelle
FAfish-audio

Fish Audio: Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

Context
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Eingabe: $100.00Ausgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
FAfish-audio

Fish Audio: S1

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

Context
Nicht angegeben
Eingabe
text
Ausgabe
speech
Eingabe: $15.00Ausgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
FAfish-audio

Fish Audio: S2 Pro

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

Context
Nicht angegeben
Eingabe
text
Ausgabe
speech
Eingabe: $15.00Ausgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
FAfish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

Context
Nicht angegeben
Eingabe
text
Ausgabe
speech
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen