FA
fish-audio

Fish Audio: Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

概要

モデル仕様

コンテキスト
情報なし
最大出力
情報なし
アーキテクチャ
audio->transcription
トークナイザー
Other
知識カットオフ
情報なし
モデレーション
いいえ
機能とモダリティ

入力出力

入力
audio
出力
transcription
API

対応APIパラメータ

利用可能なプロバイダー

1 プロバイダー

Fish Audio

unknown
コンテキスト
情報なし
最大出力
情報なし
入力
$100.00
出力
無料
キャッシュ読込
情報なし
キャッシュ書込
情報なし
fish-audio

モデル

すべてのモデル
FAfish-audio

Fish Audio: S1

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

コンテキスト
情報なし
入力
text
出力
speech
入力: $15.00出力: 無料100万トークンあたり
モデル詳細を見る
FAfish-audio

Fish Audio: S2 Pro

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

コンテキスト
情報なし
入力
text
出力
speech
入力: $15.00出力: 無料100万トークンあたり
モデル詳細を見る
FAfish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

コンテキスト
情報なし
入力
text
出力
speech
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
FAfish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

コンテキスト
情報なし
入力
text
出力
speech
入力: $15.00出力: 無料100万トークンあたり
モデル詳細を見る