モデル仕様
- コンテキスト
- 情報なし
- 最大出力
- 情報なし
- アーキテクチャ
- text->speech
- トークナイザー
- Other
- 知識カットオフ
- 情報なし
- モデレーション
- いいえ
入力 → 出力
対応APIパラメータ
1 プロバイダー
Fish Audio
unknown- コンテキスト
- 情報なし
- 最大出力
- 情報なし
- 入力
- $15.00
- 出力
- 無料
- キャッシュ読込
- 情報なし
- キャッシュ書込
- 情報なし
S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...
Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.
S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...
S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.
S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...