Google unveils Gemini 3.8 Flash TTS and Gemini 3.8 Flash‑Lite TTS, its most expressive audio generators yet, supporting over 100 languages
gemini google voice
| Source: Techmeme | Original article
Google has launched Gemini 3.8 Flash TTS and Gemini 3.8 Flash‑Lite TTS, its most expressive audio generation models yet, supporting over 100 languages.
Google rolled out two new text‑to‑speech models on 23 September 2026, branding them Gemini 3.8 Flash TTS and Gemini 3.8 Flash‑Lite TTS. In Google’s own announcement the models are described as “our most expressive audio generation models yet” and they support more than 100 languages. Both variants are now accessible through the Gemini API and Google AI Studio, positioning them for developers who need scalable, high‑quality voice output.
The Flash family builds on the Gemini 3.8 line that we covered earlier this week, when Google first introduced its Gemini 3.8 TTS capabilities. The new Flash and Flash‑Lite versions add deep creative control: developers can generate entirely new voices from natural‑language prompts, a feature aimed at gaming, immersive audiobooks, podcasts and other interactive media. Flash‑Lite is pitched as a lighter, cost‑efficient option for large‑scale dubbing and voice‑over pipelines, while Flash targets richer, character‑driven applications.
Why it matters is twofold. First, the multilingual breadth lowers barriers for global content creators, enabling rapid localisation without separate voice‑over contracts. Second, the expressiveness and speed promised by the Flash architecture could shift the economics of voice‑driven products, making custom‑voice generation viable for smaller studios and SaaS platforms. The release also nudges Google’s broader AI strategy, tying voice synthesis more tightly to its Gemini API and AI Studio ecosystem.
What to watch next includes developer uptake and pricing signals, as well as performance benchmarks against competing services from OpenAI, Microsoft and emerging European players. Observers will also be keen to see how the models integrate with Google’s upcoming AI assistant rules under the UK CMA, and whether the expressive capabilities spark new use cases in education, accessibility and interactive entertainment.
Sources
Back to AIPULSEN