モデル仕様
- コンテキスト
- 情報なし
- 最大出力
- 情報なし
- アーキテクチャ
- text->speech
- トークナイザー
- Other
- 知識カットオフ
- 情報なし
- モデレーション
- いいえ
入力 → 出力
対応APIパラメータ
1 プロバイダー
Fish Audio
unknown- コンテキスト
- 情報なし
- 最大出力
- 情報なし
- 入力
- $15.00
- 出力
- 無料
- キャッシュ読込
- 情報なし
- キャッシュ書込
- 情報なし
S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.
Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.
S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...
S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...
S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...