Tool: Gemini 3.8 TTS Playground

Simon Willison's Weblog ·

Simon Willison describes Google’s release of gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts. He writes that the models offer over 2,000 voices and custom voices from a 30-second sample of a voice the user owns or has rights to use, and describes defining conversations with different voices and voice style instructions. Read 3 viewpoints with supporting evidence and source links.

Understand this piece

3 key points

Synthesis

  1. Gemini 3.8 TTS models launched

    Google released two new Gemini text-to-speech models: gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts.

    Supporting evidence 1

    Original excerpt

    Google released two new Gemini text-to-speech models today - gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts .

    Simon Willison · Paragraph 1

    Read in source context →

    Continue exploring

    AI model release →
  2. Custom voice creation with 30-second sample

    The Gemini TTS models support custom voice creation using just a 30-second audio sample of a voice the user owns or has rights to use.

    Supporting evidence 1

    Original excerpt

    They come with a library of over 2,000 voices, plus the ability to create a custom voice with "just a 30-second audio sample of your voice or a voice you have the rights to use".

    Simon Willison · Paragraph 2

    Read in source context →

    Continue exploring

    voice customization →
  3. API supports multi-character conversations with distinct voices

    The Gemini TTS API enables defining full conversations between multiple characters, each assigned a distinct voice and voice style instructions.

    Supporting evidence 1

    Original excerpt

    A notable feature of the API is that it makes it easy to define a full conversation between multiple characters, each with different voices and voice style instructions.

    Simon Willison · Paragraph 4

    Read in source context →

    Continue exploring

    multi-character TTS →

Key passages3

Attributed passages with the context to verify them. Open the original text to check the source.

multi-character TTS

API supports multi-character conversations with distinct voices

Original excerpt

A notable feature of the API is that it makes it easy to define a full conversation between multiple characters, each with different voices and voice style instructions.
voice customization

Custom voice creation with 30-second sample

Original excerpt

They come with a library of over 2,000 voices, plus the ability to create a custom voice with "just a 30-second audio sample of your voice or a voice you have the rights to use".

Mentioned here

All mentioned things

gemini-3.8-flash-lite-tts

Mention only

Google released the gemini-3.8-flash-lite-tts model as one of two new Gemini text-to-speech models.

Read supporting evidence · Simon Willison

gemini-3.8-flash-tts

Mention only

Google released the gemini-3.8-flash-tts model as one of two new Gemini text-to-speech models.

Read supporting evidence · Simon Willison

Source & methodology

These viewpoints are linked to their original sources. Paraphrases are labeled and are not verbatim quotes.

Open transcript or source material (opens in a new tab)Report an issue

Explore these viewpoints by person