INDEX 47 ▼1 todaySPLIT OF THE DAY OpenAI annual recurring revenue approaches $70 billion56 STORIES · 565 REACTIONSANTI-AI 74% · MIDDLE GROUND 16% · PRO-AI 9%LATEST AI researchers warn superintelligence extinction risk is around 50 percent
17 sources20 reactions

Google launches Gemini 3.8 Flash TTS and Flash-Lite TTS models

55 BoomStory + reactionsProduct launch, new capability framed as progress
17 sources · newsbytesapp.com · gadgets360.com · AI News
  • Boom: Gemini 3.8 Flash TTS and Flash-Lite TTS released via API and Google AI Studio on September 23
  • Boom: Models support over 100 languages and more than 2,000 voices
  • Boom: Custom voice design from text prompts included in both models
  • Neutral: SynthID watermarking integrated to identify AI-generated audio
  • Doom: Voice replication capability included, raising potential misuse concerns
The story in full

Google released two new text-to-speech models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, on September 23, 2026. The models are available through the API and Google AI Studio, support more than 100 languages, offer over 2,000 voices, and include custom voice design from text prompts as well as SynthID watermarking.

The release also includes voice replication capability, according to reporting from Tech in Asia. The two models are positioned as distinct tiers, with Flash-Lite likely serving as a lighter, lower-cost option alongside the standard Flash variant.

Analysis

391 words

On September 23, 2026, Google released two new text-to-speech models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, through the Gemini API and Google AI Studio. The models support more than 100 languages, offer over 2,000 voices, allow custom voice design from text prompts, and include voice replication. Every clip carries SynthID watermarking, an imperceptible identifier for AI-generated audio. Google also announced the models would roll out to Google Vids and Gemini Enterprise in the near future.

The inclusion of voice replication is the detail that gives the release its sharpest edge. Voice cloning has historically been treated cautiously by major platforms, so its open availability here marks a shift in what large providers are willing to ship commercially. SynthID watermarking is Google's stated answer to misuse concerns, though watermarking is a deterrent rather than a hard barrier. The Flash and Flash-Lite tier structure also signals a push for cost efficiency, with simonw noting that most individual experiments cost less than a cent.

The Pro-AI camp is largely enthusiastic. ttul tested the model as a replacement for ElevenLabs' v3 in a live application and found the output "extremely expressive" and on par with the competitor. simonw highlighted the low cost. Google's own communications framed SynthID as a transparency feature that protects creators. Critics are less convinced. yipinwong argued Google is "spreading too thin" across too many simultaneous AI products without coherent focus, while rcr-anti pointed to inconsistent availability across Google's consumer, prosumer, and cloud platforms. nater5000 found nothing particularly impressive, and Leon Skum questioned whether any of these products generates meaningful revenue outside government contracts. The middle ground camp views the release as evidence that voice cloning has become commoditized, with simonw observing that Google's willingness to ship it reflects how widely available the capability already is elsewhere. Even so, several listeners find the output unconvincing: burkaman said results fall short of prompt accuracy, m3kw9 described the voices as over-expressive in an artificial way, and AyanamiKaine suggested that while many listeners may not notice, something present in real human speech remains absent.

The practical test will come as Google rolls out the models to Google Vids and Gemini Enterprise. Adoption rates in those products, alongside any policy response to the voice replication feature, will clarify whether Google's open approach to voice cloning becomes an industry norm or prompts regulatory pushback.

Pro-AI7

What Pro-AI voices are sayingThe models are praised for expressive, high-quality output competitive with established providers like ElevenLabs, low cost per use, and broad language and voice support, with SynthID watermarking highlighted as a transparency benefit.

Quote 1 of 7
I gave 3.8 a whirl today, replacing Eleven v3 TTS in an internal application that uses TTS to provide a listening function. The Google model produces extremely expressive output. To my ears, it’s on par with Eleven v3, which was
Anti-AI7

What Anti-AI voices are sayingCritics question Google's strategic focus, arguing the company is spreading too thin across too many AI products while lacking coherent platform alignment, and some dismiss the release as unimpressive or financially unsustainable without government contracts.

Quote 1 of 7
How long is this stored? What could go wrong? :P
accountrequiredvia Hacker News
Middle Ground6

What Middle Ground voices are sayingSome see the release as a sign that voice cloning has become commoditized enough for open availability, while a notable minority argues the voices still sound artificial and fall short of prompt accuracy despite the technological advance.

Quote 1 of 6
Seems like voice actors are safe for now. This is technologically incredible, but the results are really not very good, and usually not particularly close to the prompt.

Add your take

0 reader votes

Sign in with Google to pick a side and post. Your vote moves the story's Doom / Boom score.

More Pro-AI reactions (6)
  • “it's very expensive - most of my experiments have cost less than a cent.”

    simonw, Hacker News · 17:19 UTC
  • “Would be great if this would power the Google Books app feature. The voice system there is pretty out of date.”

    xnx, Hacker News · 16:32 UTC
  • “プログラミングや複雑な推論が強化されたこのモデル、皆さんはどんな作業に使ってみたいですか?🤖”

    トモキ|AI副業・自動化ラボ, Bluesky · 09:35 UTC
  • “To help protect creators and keep AI speech transparent, every AI-generated audio clip includes imperceptible SynthID watermarking.”

    Google {bot}, Bluesky · 15:30 UTC
  • “Google dévoile une avancée marquante dans l'IA vocale avec sa technologie Gemini 3.8 Live, promettant des interactions plus fluides.”

    iatechauquotidien.bsky.social, Bluesky · 12:35 UTC
  • “These new Gemini Audio models are rolling out today across @GoogleAIStudio, @Gemini_Notebook, and the Gemini API. Coming soon to Google Vids and Gemini Enterprise.”

    Google {bot}, Bluesky · 15:25 UTC
More Anti-AI reactions (6)
  • “Pet peeve on Google's AI rollouts: there's no alignment across the three platforms they have, consumer, prosumer, cloud. Scroll to the end of every release, including this one, and you'll see different availabilities. The fun part is the models don't”

    rcr-anti, Hacker News · 18:22 UTC
  • “Google is spreading too thin, as gemini isn't really that intelligent. They are creating gemini SOTA (not really any more), flash versions, text-to-speech, video (omni), etc. I can see they want to create an ecosystem, but I see no focus”

    yipinwong, Hacker News, skeptic · 19:26 UTC
  • “Nothing stands out is being particularly interesting or impressive about this.”

    nater5000, Hacker News · 16:53 UTC
  • “she'd "had google gemini make the sheet" for her. The changes she needed help with were... deleting a row and duplicating the sheet.”

    Lonicera (RE era), Bluesky · 11:47 UTC
  • “Google assistant is officially dead and now I gotta use Gemini”

    🍚 THICKYrice 🍚, Bluesky · 15:51 UTC
  • “None of them are making money 💰 from AI without government contracts.”

    Leon Skum, Bluesky · 12:14 UTC
More Middle Ground reactions (5)
  • “Still sounds AI, you can tell they exaggerate all the tone and trailing "high scoring expressive sounds" like your job depends on it.”

    m3kw9, Hacker News, skeptic · 16:57 UTC
  • “I love that the state of the industry is such that Google can do this and release it publicly, because that means eventually an equivalent product can come from someone else and be used locally / confidently that the generated”

    Multicomp, Hacker News · 16:24 UTC
  • “I guess voice cloning is widely enough available now from other providers that Google are no longer hesitant to ship it.”

    simonw, Hacker News · 16:19 UTC
  • “Still, all voices sound like they are missing something only real human speech can sound like. But many people will not notice the difference between AI and normal voices.”

    AyanamiKaine, Hacker News · 19:06 UTC
  • “"Super tinny monotone robotic voice" does not sound neither tinny nor monotone. Compared to what TTS from 90s sounded like.”

    112233, Hacker News, skeptic · 16:23 UTC
Pro-AI 7 · Anti-AI 7 · Middle Ground 60 reader takes

Sources

18 articles from 17 outlets
  1. newsbytesapp.comGoogle launches Gemini 3.8 Flash TTS models with voice cloning
  2. gadgets360.comGoogle Launches Gemini 3.8 Flash TTS Models With Custom Voice Creation and Control
  3. AI NewsGoogle launches Gemini 3.8 Flash TTS voice models
  4. FoneArena.comGoogle rolls out Gemini 3.8 Flash TTS and Flash-Lite TTS with custom voice creation, 2000+ voices, 100+ language support
  5. blockchain.newsGoogle Unveils Gemini 3.8 Flash TTS Models for Dynamic Voice Creation
  6. shattered.ioGemini 3.8 Flash TTS Lets You Design AI Voices From Text
  7. newsbytesapp.comGoogle's new AI models can create voices from text prompts
  8. newskarnataka.comGoogle unveils Gemini voice models for more expressive AI audio
  9. GIGAZINEGoogle releases speech synthesis AI 'Gemini 3.8 Flash TTS' and 'Gemini 3.8 Flash-Lite TTS'
  10. Tech in AsiaGoogle launches Gemini TTS with voice replication
  11. LatestLYGemini 3.8 Flash TTS, Gemini 3.8 Flash-Lite TTS Introduced by Google With Custom Voice Design and SynthID
  12. TradingViewGoogle Introduces Two New Text-To-Speech Models: Gemini 3.8 Flash TTS & Gemini 3.8 Flash-Lite TTS
  13. Unite.AIGoogle Rolls Out Gemini 3.8 Speech Models In API And AI Studio
  14. Crypto BriefingGoogle unveils Gemini 3.8 text-to-speech models for expressive, multilingual voices
  15. Google DeepMindGemini 3.8 text-to-speech says hello
  16. Hacker News front page (AI)Gemini 3.8 text-to-speech says hello
  17. blog.googleGemini 3.8 text-to-speech says hello
  18. Google DeepMind blogGemini 3.8 text-to-speech says hello