Google has announced the launch of two new text-to-speech models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, designed to transform voice generation into a dynamic creative studio. These models allow creators, developers, and enterprises to create more expressive audio experiences with improved user experiences in products like Gemini Notebook and Google Vids.

Gemini 3.8 Flash TTS is built for deep creative direction and character design, enabling users to create entirely new voices from scratch using natural language prompts. This model provides granular control over acting cues, pacing, dialect shifts, and backchanneling, making it ideal for gaming, immersive audiobooks, podcasts, and interactive media.

Gemini 3.8 Flash-Lite TTS, on the other hand, is optimized for high-volume, cost-efficient scale, making it suitable for high-volume dubbing, audio content creation, and expressive voice agents. Both models complement Google's fast-growing Gemini Audio family, which includes 3.5 Live Translate, 3.5 Transcribe, 3.8 Live, and 3.8 Live Extended Thinking.

With Gemini 3.8 Flash TTS, users can create and customize their own voices, scaling up from 30 original voices to an infinite library. The model also enables generative voice design, allowing users to create bespoke voices from scratch by customizing role, accent, and voice characteristics across over 100 languages and dialects using natural language prompting.

The Gemini 3.8 Flash TTS model has secured the #1 overall spot on Hume AI's Voice Design Benchmark and also leads in accent modeling. In blind human preference evaluations on Voice Arena, Gemini 3.8 Flash and Flash-Lite TTS have secured top positions amongst competitors in key global languages, including Japanese, Brazilian Portuguese, Vietnamese, Modern Standard Arabic, Mexican Spanish, and Hindi.

Google has built its voice creation and replication capabilities with strict safeguards to protect voice talent, respect identity, and ensure content transparency. The company has also introduced a new audio playground, Google AI Studio, where developers can experience the new speech generation capabilities and deploy high-performance voice interfaces with ease.

Este artigo foi escrito com a assistência de IA.
News Factory APP - notícias agênticas para impulsionar seu SEO e AEO.