Todos los modelos
generación de audio
Vista previaGemini 3.1 Flash TTS
Gemini 3.1 Flash TTS is Google’s preview text-to-speech model for controllable audio from exact text. It detects the spoken language automatically across more than 70 languages, and Google documents single-speaker and two-speaker output, natural-language direction for style, accent, pace, and tone, and 30 prebuilt voices; Gobostock keeps each voice tied to the engine it belongs to.
The audio tool calculates the final credit quote from the selected operation and output settings. El presupuesto que se muestra en la herramienta es el que pagas.
0:000:06
Audio en pausa
Ideal para
- Narration that must recite exact text
- Designed voices using Google’s prebuilt speaker catalog
Límites conocidos
- The Gemini TTS endpoint is a preview model.
- Generated voices must remain bound to the Google runtime used to create them.
Capacidades
- Single-speaker and two-speaker text-to-speech
- Natural-language control over style, accent, pace, and tone
- 30 prebuilt Google voices with automatic language detection across 70+ languages
Entradas
- Exact script
- Prebuilt voice selection
- Optional style and delivery direction
Salidas
- Generated speech audio
Hecho en Gobostock
Más de Gemini 3.1 Flash TTS
0:000:05
Audio en pausa
0:000:04
Audio en pausa
0:000:06
Audio en pausa
0:000:05
Audio en pausa
Vistas previas con marca de agua
Comparar