FA
fish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

模型概览

模型规格

上下文
暂未提供
最大输出
暂未提供
架构
text->speech
分词器
Other
知识截止时间
暂未提供
内容审核
能力与模态

输入输出

输入
text
输出
speech
API

支持的 API 参数

可用提供方

1 提供方

Fish Audio

unknown
上下文
暂未提供
最大输出
暂未提供
输入
$15.00
输出
免费
缓存读取
暂未提供
缓存写入
暂未提供
fish-audio

个模型

全部模型
FAfish-audio

Fish Audio: Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

上下文
暂未提供
输入
audio
输出
transcription
输入: $100.00输出: 免费每百万 Token
查看模型详情
FAfish-audio

Fish Audio: S1

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

上下文
暂未提供
输入
text
输出
speech
输入: $15.00输出: 免费每百万 Token
查看模型详情
FAfish-audio

Fish Audio: S2 Pro

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

上下文
暂未提供
输入
text
输出
speech
输入: $15.00输出: 免费每百万 Token
查看模型详情
FAfish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

上下文
暂未提供
输入
text
输出
speech
输入: 免费输出: 免费每百万 Token
查看模型详情