• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
Technology
Announcement

Google launches Gemini 3.8 Flash TTS and Flash-Lite TTS

Google says Flash TTS can create custom voices from text prompts across more than 100 languages and dialects, while Flash-Lite targets high-volume dubbing, audio creation and voice agents.

GoogleGO
Google AI StudioGA
🚨 AI News | TestingCatalog🚨A
13 Sources, 17d ago, first seen 17d ago

TLDR

Google announced the two text-to-speech models on September 23, 2026, with rollout beginning that day in the Gemini API and Google AI Studio. The company lists a library of 2,000-plus voices and says both models support line-by-line performance direction and two-speaker scenes. Voice remixing was listed as coming soon.

Google says voice replication uses a 30-second sample of your own voice or one you have rights to use, and requires a matching verbal consent recording from the voice owner. It says generated audio carries SynthID watermarks. Unite.AI reports that voice replication through AI Studio is unavailable in Illinois, Texas, the European Economic Area, the UK, Switzerland and India.

Combined views

2.8M

13 Sources, first seen 17d ago

11.3K likes806 comments3.2K saves958 reposts

Combined views

2.8M

13 Sources, first seen 17d ago

11.3K likes806 comments3.2K saves958 reposts
Google launch graphic introducing Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS on a pale blue background.
Image: Source: Google

Google has released two Gemini text-to-speech models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, for developers using the Gemini API and Google AI Studio. In its announcement, the company describes Flash TTS as the higher-capability model and Flash-Lite as the lower-cost option for high-volume work such as dubbing and voice agents.

Featured Source

Both models are built around instructions for delivery, not just the words to be spoken. Google says developers can direct a line’s pace, tone and style, and can create two-speaker scenes. The company also says the release supports more than 100 languages and dialects, a voice-design workflow based on text prompts, and a catalog of more than 2,000 production-ready voices.

What the two models are for

The distinction is primarily about the trade-off between expressive output and scale. Flash TTS is intended for more nuanced delivery; Flash-Lite TTS is positioned for applications that need to generate a large amount of audio efficiently. Those are product claims from Google, rather than an independent comparison of quality or cost.

Google is also introducing voice replication. Its announcement says that feature uses a 30-second sample and requires a matching verbal-consent recording from the voice owner. The company says users must have the right to use the voice and that generated audio carries SynthID watermarking.

Availability still has boundaries

The models began rolling out on Sept. 23 through the Gemini API and Google AI Studio. Google lists voice remixing as a future feature, not part of the initial release. Availability can also vary by product and location: Unite.AI reported that voice replication in AI Studio is unavailable in Illinois, Texas, the European Economic Area, the UK, Switzerland and India.

For developers, the immediate change is a new pair of API-accessible speech models with more controls over how a line is performed. The broader questions, including how well those controls hold up across languages and real production workloads, will require testing beyond Google’s launch materials.

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Related

Google unveils universal Gemini agent for enterprise

Google says the agent can create sub-agents for multi-step tasks and maintain context across devices.

Google is reportedly betting Gemini can become the intelligence layer for many robots

The Information says Google DeepMind CEO Koray Kavukcuoglu discussed the company’s robotics ambitions in an interview.

Google AI Pro offers access to a 24/7 personal AI agent in Gemini

A post reshared by Google pitches the agent as a way to delegate routine chores when unread emails, meeting requests and to-dos pile up.

15 Sources

GoogleGemini 3.8 text-to-speech says hello17d
Unite.AIGoogle Rolls Out Gemini 3.8 Speech Models In API And AI Studio17d
Google@GoogleCan you hear that? Our Gemini Audio family is getting louder 🔊 We're introducing two of our most expressive audio generation models yet from @GoogleDeepMind: Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS.17d
Google AI Studio@GoogleAIStudiointroducing Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, our most expressive audio generation models yet these models enable creators, developers, and enterprises to create richer, more expressive audio experiences try them via the Gemini API and in AI Studio: https://aistudio.google.com/generate-speech?model=gemini-3.8-flash-tts&e=017d
🚨 AI News | TestingCatalog@testingcatalogGOOGLE 🔥: Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS are now live on Google AI Studio and APIs! > "Flagship TTS model for Voice Design and dual-speaker screenplay control. Prompt custom vocal personas, direct line-by-line delivery, and add vocal bursts." > "Our most expressive audio generation models yet. TTeSting time! 👀17d
Logan Kilpatrick@OfficialLoganKIntroducing Gemini 3.8 Flash and Flash-Lite TTS, our new SOTA text to speech model with: - a new voice design experience - 2,000+ production ready voices - voice replication - support for 100 languages - voice remixing (soon) - #1 spot on Hume AI's voice benchmarks and more!!17d
Techmeme@TechmemeGoogle releases Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, its "most expressive audio generation models yet", with support for more than 100 languages (Google) (Visit Techmeme dot com for the link and full context!)17d
SiliconANGLE@SiliconANGLEGoogle launches two benchmark-topping speech generation models https://ift.tt/wzpdfWL16d
Android Authority@AndroidAuthGemini can now clone your voice and perform scripts like an actor https://www.androidauthority.com/gemini-3-8-flash-text-to-speech-rolling-out-3714915/16d
The New Stack@thenewstackGoogle's Gemini 3.8 TTS lets developers design a voice from a text prompt or replicate one from a short clip, then reuse it by ID through the Gemini API. https://thenewstack.io/gemini-tts-voice-replication-api/?taid=6ab52cead1581a0001748df3&utm_campaign=trueanthem&utm_medium=social&utm_source=twitter16d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    GeminiGoogleGemini 3.8 Flash
    Gemini 3.8 Flash TTS
    GoogleDeepMind
    Logan Kilpatrick

    15 Sources

    GoogleGemini 3.8 text-to-speech says hello17d
    Unite.AIGoogle Rolls Out Gemini 3.8 Speech Models In API And AI Studio17d
    Google@GoogleCan you hear that? Our Gemini Audio family is getting louder 🔊 We're introducing two of our most expressive audio generation models yet from @GoogleDeepMind: Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS.17d
    Google AI Studio@GoogleAIStudiointroducing Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, our most expressive audio generation models yet these models enable creators, developers, and enterprises to create richer, more expressive audio experiences try them via the Gemini API and in AI Studio: https://aistudio.google.com/generate-speech?model=gemini-3.8-flash-tts&e=017d
    🚨 AI News | TestingCatalog@testingcatalogGOOGLE 🔥: Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS are now live on Google AI Studio and APIs! > "Flagship TTS model for Voice Design and dual-speaker screenplay control. Prompt custom vocal personas, direct line-by-line delivery, and add vocal bursts." > "Our most expressive audio generation models yet. TTeSting time! 👀17d
    Logan Kilpatrick@OfficialLoganKIntroducing Gemini 3.8 Flash and Flash-Lite TTS, our new SOTA text to speech model with: - a new voice design experience - 2,000+ production ready voices - voice replication - support for 100 languages - voice remixing (soon) - #1 spot on Hume AI's voice benchmarks and more!!17d
    Techmeme@TechmemeGoogle releases Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, its "most expressive audio generation models yet", with support for more than 100 languages (Google) (Visit Techmeme dot com for the link and full context!)17d
    SiliconANGLE@SiliconANGLEGoogle launches two benchmark-topping speech generation models https://ift.tt/wzpdfWL16d
    Android Authority@AndroidAuthGemini can now clone your voice and perform scripts like an actor https://www.androidauthority.com/gemini-3-8-flash-text-to-speech-rolling-out-3714915/16d
    The New Stack@thenewstackGoogle's Gemini 3.8 TTS lets developers design a voice from a text prompt or replicate one from a short clip, then reuse it by ID through the Gemini API. https://thenewstack.io/gemini-tts-voice-replication-api/?taid=6ab52cead1581a0001748df3&utm_campaign=trueanthem&utm_medium=social&utm_source=twitter16d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet