Tool: Gemini 3.8 TTS Playground

Simon Willison's Weblog ·

Simon Willison describes Google’s release of gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts. He writes that the models offer over 2,000 voices and custom voices from a 30-second sample of a voice the user owns or has rights to use, and describes defining conversations with different voices and voice style instructions. Lee 3 puntos de vista con sus evidencias y enlaces a las fuentes.

De un vistazo

  • Gemini 3.8 TTS models launched

    Google released two new Gemini text-to-speech models: gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts.

    Ver el momento de apoyo · Párrafo 1
  • Custom voice creation with 30-second sample

    The Gemini TTS models support custom voice creation using just a 30-second audio sample of a voice the user owns or has rights to use.

    Ver el momento de apoyo · Párrafo 2
  • API supports multi-character conversations with distinct voices

    The Gemini TTS API enables defining full conversations between multiple characters, each assigned a distinct voice and voice style instructions.

    Ver el momento de apoyo · Párrafo 4

Pasajes clave3

Pasajes atribuidos con contexto para verificarlos. Abra el texto original para comprobar la fuente.

multi-character TTS

API supports multi-character conversations with distinct voices

Extracto original

A notable feature of the API is that it makes it easy to define a full conversation between multiple characters, each with different voices and voice style instructions.
voice customization

Custom voice creation with 30-second sample

Extracto original

They come with a library of over 2,000 voices, plus the ability to create a custom voice with "just a 30-second audio sample of your voice or a voice you have the rights to use".

Fuente y metodología

Estas perspectivas enlazan a sus fuentes originales. Las paráfrasis están identificadas y no son citas textuales.

Abrir transcripción o material de origen (se abre en una pestaña nueva)Reportar un problema