Todos los modelos
generación de audio
Fish Audio S2.1 Pro
Fish Audio S2.1 Pro is a multilingual text-to-speech model for expressive, low-latency narration and dialogue. It supports more than 80 languages, automatic language detection, inline emotion and delivery cues, multi-speaker dialogue, voice-library choices, and supported reference-voice cloning.
The audio tool calculates the final credit quote from the selected operation and output settings. El presupuesto que se muestra en la herramienta es el que pagas.
Todavía no se ha publicado ningún ejemplo.
Ideal para
- Expressive narration and character dialogue
- Multilingual speech with automatic language detection
Límites conocidos
- The Fish preset path is retained internally and may not appear in every public creator workflow.
- Voice cloning requires rights and consent for the reference speaker.
Capacidades
- Expressive multilingual text-to-speech
- Inline emotion cues and two-speaker dialogue
- Voice-library and supported reference-voice workflows
Entradas
- Script
- Voice or supported reference voice
- Speed, volume, and expression settings
Salidas
- Generated speech audio
Comparar