すべてのモデル
オーディオ生成
Fish Audio S2.1 Pro
Fish Audio S2.1 Pro is a multilingual text-to-speech model for expressive, low-latency narration and dialogue. It supports more than 80 languages, automatic language detection, inline emotion and delivery cues, multi-speaker dialogue, voice-library choices, and supported reference-voice cloning.
The audio tool calculates the final credit quote from the selected operation and output settings. ツールに表示される見積もりが、お支払いいただく金額です。
まだサンプルは公開されていません。
最適な用途
- Expressive narration and character dialogue
- Multilingual speech with automatic language detection
既知の制限
- The Fish preset path is retained internally and may not appear in every public creator workflow.
- Voice cloning requires rights and consent for the reference speaker.
機能
- Expressive multilingual text-to-speech
- Inline emotion cues and two-speaker dialogue
- Voice-library and supported reference-voice workflows
入力
- Script
- Voice or supported reference voice
- Speed, volume, and expression settings
出力
- Generated speech audio
比較