Todos los modelos

generación de audio

Vista previa

Gemini 3.1 Flash TTS

Gemini 3.1 Flash TTS is Google’s preview text-to-speech model for controllable audio from exact text. It detects the spoken language automatically across more than 70 languages, and Google documents single-speaker and two-speaker output, natural-language direction for style, accent, pace, and tone, and 30 prebuilt voices; Gobostock keeps each voice tied to the engine it belongs to.

The audio tool calculates the final credit quote from the selected operation and output settings. El presupuesto que se muestra en la herramienta es el que pagas.

0:000:06
Audio en pausa

Ideal para

  • Narration that must recite exact text
  • Designed voices using Google’s prebuilt speaker catalog

Límites conocidos

  • The Gemini TTS endpoint is a preview model.
  • Generated voices must remain bound to the Google runtime used to create them.

Capacidades

  • Single-speaker and two-speaker text-to-speech
  • Natural-language control over style, accent, pace, and tone
  • 30 prebuilt Google voices with automatic language detection across 70+ languages

Entradas

  • Exact script
  • Prebuilt voice selection
  • Optional style and delivery direction

Salidas

  • Generated speech audio

Hecho en Gobostock

Más de Gemini 3.1 Flash TTS

0:000:05
Audio en pausa
0:000:04
Audio en pausa
0:000:06
Audio en pausa
0:000:05
Audio en pausa

Vistas previas con marca de agua

Comparar

Los nombres de los modelos son marcas comerciales de sus respectivos propietarios. Gobostock no está afiliado a ellos ni cuenta con su respaldo, salvo que se indique lo contrario.