模型规格
- 上下文
- 暂未提供
- 最大输出
- 暂未提供
- 架构
- text->speech
- 分词器
- Other
- 知识截止时间
- 暂未提供
- 内容审核
- 否
输入 → 输出
支持的 API 参数
1 提供方
Fish Audio
unknown- 上下文
- 暂未提供
- 最大输出
- 暂未提供
- 输入
- $15.00
- 输出
- 免费
- 缓存读取
- 暂未提供
- 缓存写入
- 暂未提供
S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...
Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.
S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.
S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...
S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...