Google unveils Gemini 3.8 Flash TTS with advanced voice customization

1 hour ago 1



Google just made it a lot harder to tell whether the voice reading your audiobook belongs to a human or a very well-trained language model. The company launched its Gemini 3.8 Flash TTS models on September 23, giving developers an expansive toolkit for generating synthetic speech that sounds less like a GPS navigator and more like an actual person with opinions about inflection. The release includes two model variants: gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts. Both are available through the Gemini API and Google AI Studio, positioning them squarely at developers building voice-driven applications at scale. What the new models actually do At the core of the 3.8 Flash TTS lineup is a layered approach to voice selection. Users get three tiers to work with: 30 prebuilt studio voices for quick deployment, an Extended Voice Library with hundreds of additional options, and a custom voice design system that lets you describe the voice you want using natural language prompts. That last feature is worth pausing on. Instead of tweaking sliders for pitch and speed like some 2019-era audio editor, developers can write something closer to a character description. The system then genera...

Read Entire Article