By Maria Deutscher
Publication Date: 2026-09-24 00:22:00
Google LLC today made two new text-to-speech models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, available through its cloud platform.
The algorithms have highly similar application programming interfaces, which makes using them side-by-side relatively simple for developers. Flash-Lite TTS is optimized for cost-cost efficiency and inference speed. Flash TTS offers better audio quality for a higher price. Google envisions customers using it for tasks such as creating audiobooks.
There are also other differences between the models. Most notably, Flash TTS can generate speech in 130 languages on launch while Flash-Lite TTS supports 101.
Both models offer access to a library of more than 2,000 prepackaged voices. Developers can create custom voices with natural language prompts. Google makes it possible to customize parameters such an AI speaker’s vocal timbre, accent and pacing.
The second way to customize the new models is to generate a synthetic voice based on a 30-second…


