すべてのモデル

オーディオ生成

Fish Audio S2.1 Pro

Fish Audio S2.1 Pro is a multilingual text-to-speech model for expressive, low-latency narration and dialogue. It supports more than 80 languages, automatic language detection, inline emotion and delivery cues, multi-speaker dialogue, voice-library choices, and supported reference-voice cloning.

The audio tool calculates the final credit quote from the selected operation and output settings. ツールに表示される見積もりが、お支払いいただく金額です。

まだサンプルは公開されていません。

最適な用途

  • Expressive narration and character dialogue
  • Multilingual speech with automatic language detection

既知の制限

  • The Fish preset path is retained internally and may not appear in every public creator workflow.
  • Voice cloning requires rights and consent for the reference speaker.

機能

  • Expressive multilingual text-to-speech
  • Inline emotion cues and two-speaker dialogue
  • Voice-library and supported reference-voice workflows

入力

  • Script
  • Voice or supported reference voice
  • Speed, volume, and expression settings

出力

  • Generated speech audio

比較

モデル名はそれぞれの所有者の商標です。特に記載がない限り、Gobostockはそれらの所有者と提携しておらず、その推奨を受けているものでもありません。