From capability to creation · Character voices
Give a character a voice, then direct each line: Gemini 3.8 TTS
One knight can have different voices; a monologue can carry directions for each line. Official demos show what creators can control when voicing a short film.
Released September 23; reviewed September 24. These are Google demos. We checked the visuals and playback access, but have not generated our own audio or tested Chinese delivery.
Design the voice separately from the character
In the game demo, an armored knight stands beside a voice description and playback button. At 48 seconds, the description asks for a deep, gravelly warrior voice. At 73 seconds, it asks for a high-pitched, regal princess voice while retaining the knight. Open the video to hear those two moments.
The useful creative idea is that appearance and voice can be adjusted separately. If you already have a character or animated shot, voice becomes another way to shape personality beyond choosing a fixed preset.
- Watch and listen: one knight, two voice descriptionsAbout 84 seconds; compare 48 and 73 seconds.

Sources and further reading
- Official game-character voice demoOfficial example · About 84 secondsThe character stays while the voice description changes. Our image is the original frame at 48 seconds.
After voice design comes the performance
A second official demo divides a detective’s monologue into performance beats. Emotion and volume instructions sit beside the dialogue, with sighs, short pauses and breaths inside it. This shows two layers: who is speaking, then how this particular line is delivered.
Google launched Flash TTS and Flash-Lite TTS on September 23. Flash emphasizes character voice design; Lite targets high-volume generation. Both support line-level direction. The examples show an approach; natural Chinese emotion and consistency across repeated generations still need comparable samples and real projects.
- Watch and listen: directions inside a monologueAbout 63 seconds; an abbreviated official demo.

Sources and further reading
- Google: Gemini 3.8 TTS releaseOfficial announcement · September 23Voice design, line-level direction and rollout access.
- Official detective-monologue demoOfficial example · About 63 secondsScript beats, performance directions and pause cues; the video is shortened for illustration.
Find an entry point and bring a short scene
At launch, both models begin rolling out in AI Studio and the Gemini API. Product distribution pairs Flash with Gemini Notebook and Lite with Google Vids. This does not mean the regular Gemini chat window has every demonstrated control; check the product and account you use.
Try two or three lines of your own dialogue. Keep the script and character fixed, change one pause or emotion, and listen for whether it serves the scene. When returning to our AI film cases, notice how sound helps establish a character. This is a creative direction to explore, not a claim that those older films used the new model.
- Open speech generation in AI StudioSign-in required; availability depends on your account.
- Return to AI film projects
Sources and further reading
- Google: Gemini 3.8 TTS releaseOfficial announcement · September 23Voice design, line-level direction and rollout access.
- Gemini 3.8 Audio model cardOfficial model card · September editionDistinguishes TTS from live conversation models and lists distribution channels.