INDEX 47 ▼1 todaySPLIT OF THE DAY OpenAI annual recurring revenue approaches $70 billion56 STORIES · 565 REACTIONSANTI-AI 74% · MIDDLE GROUND 16% · PRO-AI 9%LATEST AI researchers warn superintelligence extinction risk is around 50 percent
6 sources2 reactions

Google releases two Gemini text-to-speech models with voice design

64 BoomStory toneCapability launch, framed as a product expansion
6 sources · The Tech Outlook · MarkTechPost · Google News
  • Boom: Flash TTS can generate entirely new voices from plain text descriptions
  • Doom: Voice cloning feature builds a voice profile from a 30-second audio sample
  • Boom: Both models support over 100 languages and two-voice dialogue from one script
  • Boom: Users can attach per-line stage directions to control delivery in both models
  • Neutral: Flash-Lite TTS is a lighter version without prompt-based voice creation
The story in full

Google introduced two new text-to-speech models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, on September 23, 2026. Both models support more than 100 languages, allow users to add stage directions to individual lines, and can generate two-voice dialogue from a single script. A voice cloning feature builds a voice profile from a 30-second audio sample.

The Flash TTS model adds the ability to create entirely new voices from text descriptions, a capability the Flash-Lite version does not include. The two models differ in that designation, with Flash-Lite positioned as a lighter alternative. No pricing, availability dates, or access details were provided in the sources.

Analysis

390 words

On September 23, 2026, Google announced two new text-to-speech models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS. Both support more than 100 languages, allow per-line stage directions to shape delivery, and can generate two-voice dialogue from a single script. The Flash TTS model goes further by letting users create entirely new voices from plain text descriptions, a feature absent from the lighter Flash-Lite version. Both models include a voice cloning capability that builds a voice profile from a 30-second audio sample. Google released no pricing, availability dates, or access details alongside the announcement.

The combination of features marks a meaningful step beyond conventional text-to-speech tools. Generating a voice from a text description, rather than recording or selecting from a preset library, moves voice creation into a design space where no audio input is required at all. The 30-second cloning feature further lowers the barrier to reproducing a real person's voice. Taken together, these capabilities widen access to synthetic voice production considerably, which is precisely what puts the announcement at the center of an ongoing argument about creative opportunity versus misuse potential. What remains genuinely unresolved is how Google intends to govern cloning and synthetic voice design, since the announcement included no details on safeguards, consent mechanisms, or usage restrictions.

The Pro-AI camp is framing the release in enthusiastic terms. GadgetBond.com described the new models as turning AI voice generation into something closer to a virtual voice studio, emphasizing the creative and practical reach of the toolset. Marie Haynes highlighted the voice replication feature directly, noting that users can now replicate their own voice in Google AI Studio, which signals perceived value for personal and professional workflows. The Anti-AI camp has not yet published reactions to this specific release, though stories involving voice cloning and synthetic voice creation from minimal audio samples typically draw concern from that camp around consent, impersonation risk, and the potential for audio deepfakes. The Middle Ground camp has also not weighed in yet, but stories of this kind usually prompt that camp to acknowledge the creative utility while calling for clearer platform-level rules before wide deployment.

The details most worth watching are any governance or policy terms Google attaches to access when it does announce availability, along with whether the voice cloning feature includes any verification or consent layer for the identity being cloned.

Pro-AI2
Quote 1 of 2
You can now replicate your voice in Google AI Studio.
Marie Haynesvia Bluesky
Anti-AI
No Anti-AI voice has weighed in yet. Silence is a signal too.
Middle Ground
No Middle Ground take collected yet.

Add your take

0 reader votes

Sign in with Google to pick a side and post. Your vote moves the story's Doom / Boom score.

More Pro-AI reactions (1)
  • “Google launches Gemini 3.8 Flash TTS and Flash-Lite TTS Google’s latest Gemini models turn AI voice generation into something closer to a virtual voice studio.”

    GadgetBond.com, Bluesky · 17:37 UTC
No more Anti-AI reactions
No more Middle Ground reactions
Pro-AI 2 · Anti-AI 0 · Middle Ground 00 reader takes

Sources

6 articles from 6 outlets
  1. The Tech OutlookGoogle introduces two new text-to-speech Gemini models: Gemini 3.8 Flash TTS and 3.8 Flash Lite TTS
  2. MarkTechPostGoogle Releases Gemini 3.8 Flash TTS and Flash-Lite TTS With Prompt-Based Voice Design
  3. Google NewsGoogle's new Flash TTS models let you design AI voices from scratch using text descriptions
  4. The DecoderGoogle's new Flash TTS models let you design AI voices from scratch using text descriptions
  5. Crypto BriefingGoogle unveils Gemini 3.8 Flash TTS with advanced voice customization
  6. Seeking AlphaGoogle unveils new Gemini 3.8 Flash TTS, Gemini 3.8 Flash-Lite TTS AI models (GOOG:NASDAQ)