API supports multi-character conversations with distinct voices
Originalauszug
A notable feature of the API is that it makes it easy to define a full conversation between multiple characters, each with different voices and voice style instructions.
Simon Willison's Weblog ·
Simon Willison describes Google’s release of gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts. He writes that the models offer over 2,000 voices and custom voices from a 30-second sample of a voice the user owns or has rights to use, and describes defining conversations with different voices and voice style instructions. Lies 3 Standpunkte mit Belegen und Links zu den Originalquellen.
Google released two new Gemini text-to-speech models: gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts.
Unterstützendes Moment lesen · Absatz 1The Gemini TTS models support custom voice creation using just a 30-second audio sample of a voice the user owns or has rights to use.
Unterstützendes Moment lesen · Absatz 2The Gemini TTS API enables defining full conversations between multiple characters, each assigned a distinct voice and voice style instructions.
Unterstützendes Moment lesen · Absatz 4Zugeordnete Passagen mit dem Kontext zur Überprüfung. Öffnen Sie den Originaltext, um die Quelle zu prüfen.
Originalauszug
A notable feature of the API is that it makes it easy to define a full conversation between multiple characters, each with different voices and voice style instructions.
Originalauszug
They come with a library of over 2,000 voices, plus the ability to create a custom voice with "just a 30-second audio sample of your voice or a voice you have the rights to use".
Originalauszug
Google released two new Gemini text-to-speech models today - gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts .
Diese Standpunkte sind mit ihren Originalquellen verknüpft. Paraphrasen sind gekennzeichnet und keine wörtlichen Zitate.
Transkript oder Quellenmaterial öffnen (wird in einem neuen Tab geöffnet)Ein Problem melden