Google’s new Gemini TTS models can clone a voice from 30 seconds of audio

Wait 5 sec.

“Today, we’re introducing two new text-to-speech models to the Gemini family, transforming voice generation from static presets into a dynamic creative studio,” Google wrote in its announcement. Google released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS on Wednesday. Both are rolling out from today in the Gemini API and Google AI Studio. They […]This story continues at The Next Web