模型规格
- 上下文
- 暂未提供
- 最大输出
- 暂未提供
- 架构
- text->speech
- 分词器
- Other
- 知识截止时间
- 暂未提供
- 内容审核
- 否
输入 → 输出
支持的 API 参数
1 提供方
Fish Audio
unknown- 上下文
- 暂未提供
- 最大输出
- 暂未提供
- 输入
- $15.00
- 输出
- 免费
- 缓存读取
- 暂未提供
- 缓存写入
- 暂未提供
S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.
Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.
S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...
S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...
S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...