fish-audio logo
fish-audio

Fish Audio: S1

来源介绍(英文)

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

模型概览

模型规格

上下文
暂未提供
最大输出
暂未提供
架构
text->speech
分词器
Other
知识截止时间
暂未提供
内容审核
OPENROUTER

完整价格

同步自 OpenRouter。Token 价格均按每百万 Token 展示。

UTF-8 字节
$15
每百万 UTF-8 字节
API

快速调用

在本地设置 OPENROUTER_API_KEY。Python 需安装 requests;JavaScript 在 Node.js 中运行。密钥应保存在服务端。

请选择此模型支持的声音;若未列出声音,请替换 YOUR_VOICE_ID。

API 文档

curl --fail-with-body https://openrouter.ai/api/v1/audio/speech \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"fish-audio/s1","input":"Hello!","voice":"YOUR_VOICE_ID","response_format":"mp3"}' \
  --output speech.mp3
能力与模态

输入输出

输入
text
输出
speech
API

支持的 API 参数

可用提供方

1 提供方

核验日期: 2026年9月7日

在 OpenRouter 查看实时 Provider

Provider 可用性、延迟、吞吐量和路由会持续变化,请前往来源页面查看实时运行数据。

OpenRouter

Fish Audio

暂未提供
上下文
暂未提供
最大输出
暂未提供
fish-audio

4 个模型

全部模型
fish-audio logofish-audio

Fish Audio: Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

上下文
暂未提供
输入
audio
输出
transcription
音频时长: $0.0001 每秒
查看模型详情
fish-audio logofish-audio

Fish Audio: S2 Pro

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

上下文
暂未提供
输入
text
输出
speech
UTF-8 字节: $15 每百万 UTF-8 字节
查看模型详情
fish-audio logofish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

上下文
暂未提供
输入
text
输出
speech
输入: 免费 每百万 Token输出: 免费 每百万 Token
查看模型详情
fish-audio logofish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

上下文
暂未提供
输入
text
输出
speech
UTF-8 字节: $15 每百万 UTF-8 字节
查看模型详情