conv.

All stories
AIResurging · day 2

Google launches Gemini 3.8 Flash TTS with custom voice design and 100+ languages

New text-to-speech models let users design custom character voices, direct scene dialogue, and clone voices with consent verification.

What to know

  • Google released two new TTS models — Gemini 3.8 Flash TTS and Flash-Lite TTS — with custom voice design, over 100 language support, and 2,000+ ready-made voices.
  • Voice replication requires verbal consent from the voice owner as a safeguard against misuse.
  • The models rank #1 or #2 on several third-party benchmarks (Hume AI, Artificial Analysis) for quality and pronunciation, though they trail the fastest competing TTS models on raw generation speed.
  • Availability spans Google AI Studio, the Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids.

Google Developer of Gemini modelsGoogle DeepMind AI research lab behind the models@officiallogank Google AI product figureArtificial Analysis Independent AI benchmarking tracker

Google launches Gemini 3.8 Flash TTS with custom voice design and 100+ languages
x.com

How it unfolded 2 developments, newest first · click a bar or a number to jump articlesposts

Peak 24 pieces in one hour at Today, 11 AM; 48 pieces over 3 days (15 articles · 1 video · 16 posts · 16 comments) Sep 21, 11 AM — 1 piece · 1 article — Newswires 1Sep 21, 12 PM — quietSep 21, 1 PM — quietSep 21, 2 PM — quietSep 21, 3 PM — quietSep 21, 4 PM — quietSep 21, 5 PM — quietSep 21, 6 PM — quietSep 21, 7 PM — quietSep 21, 8 PM — quietSep 21, 9 PM — quietSep 21, 10 PM — quietSep 21, 11 PM — quietYesterday, 12 AM — quietYesterday, 1 AM — quietYesterday, 2 AM — quietYesterday, 3 AM — quietYesterday, 4 AM — quietYesterday, 5 AM — quietYesterday, 6 AM — quietYesterday, 7 AM — quietYesterday, 8 AM — quietYesterday, 9 AM — quietYesterday, 10 AM — quietYesterday, 11 AM — quietYesterday, 12 PM — quietYesterday, 1 PM — quietYesterday, 2 PM — quietYesterday, 3 PM — quietYesterday, 4 PM — quietYesterday, 5 PM — quietYesterday, 6 PM — quietYesterday, 7 PM — quietYesterday, 8 PM — quietYesterday, 9 PM — quietYesterday, 10 PM — quietYesterday, 11 PM — quietToday, 12 AM — quietToday, 1 AM — quietToday, 2 AM — quietToday, 3 AM — quietToday, 4 AM — quietToday, 5 AM — quietToday, 6 AM — quietToday, 7 AM — quietToday, 8 AM — quietToday, 9 AM — quietToday, 10 AM — 1 piece · 1 article — Mastodon 1Today, 11 AM — 24 pieces · 10 articles · 1 video · 11 posts · 2 comments — Newswires 10, X 10, Hacker News 3, +1 moreToday, 12 PM — 7 pieces · 1 article · 6 comments — Hacker News 6, Newswires 1Today, 1 PM — 2 pieces · 1 article · 1 post — Mastodon 1, Newswires 1Today, 2 PM — 1 piece · 1 comment — Hacker News 1Today, 3 PM — 3 pieces · 1 article · 2 comments — Hacker News 2, Google News 1Today, 4 PM — 5 pieces · 2 posts · 3 comments — Hacker News 3, Mastodon 1, X 1Today, 5 PM — quietToday, 6 PM — quietToday, 7 PM — 1 piece · 1 post — Hacker News 1Today, 8 PM — 2 pieces · 2 comments — Hacker News 2Today, 9 PM — 1 piece · 1 post — Mastodon 1Today, 10 PM — quiet 1–2
yesterdaytodaynow · 11:21 PM ET
  1. 1

    Artificial Analysis publishes independent speed and quality benchmarks

    Third-party AI benchmarking tracker Artificial Analysis reported Gemini 3.8 Flash TTS processes 44.1 characters per second (about 2.7x realtime) versus 40.2 for Flash-Lite, placing Google's models behind faster competitors like Falcon 2, while noting the Flash model debuts at #1 on its Pronunciation Robustness Benchmark and #2 on its Provider Voice Arena Leaderboard.

    “Gemini 3.8 Flash TTS processes 44.1 characters per second, compared to 40.2 characters per second for Gemini 3.8 Flash-Lite TTS, approximately 2.7x and 2.4x faster than realtime, respectively.”
    — @artificialanlys
    1. first by The Next Web, 11h ago · also Google, Simon Willison, Simon Willison's Weblog, The Decoder, MarkTechPost

      4 more headlines
    • The new Gemini 3.8 TTS models are super-cheap and can generate conversations between multiple voices (from 2,000+, or you can clone your own) - I built a little UI for it, then had Claude knock up a script where two pelicans debate moving to Pacifica Pier

      @simonwX6h ago10▲view on X ↗
    2 more of the top 3 · 30 posts in this stretch
    • Here's a video of my Emotive Audiobook Creator, KeenLore, a locally hosted web app:https://www.youtube.com/watch?v=WAeHgE94rVoNo cloud, no tokens to pay. Reads a book using a full cast of characters. Quotation attribution detection (for my novel) is at 97.2% accuracy (485/499 quotes identified and assigned correctly). The autofill of character…

      thangalinHacker News11h agoview on Hacker News ↗
    • hn100@social.lansky.name

      Gemini 3.8 text-to-speech Link: https:// blog.google/innovation-and-ai/ models-and-research/gemini-models/gemini-3-8-text-to-speech/ Discussion: https:// news.ycombinator.com/item?id=4 9817615

      hn100@social.lansky.nameMastodon9h agoview on Mastodon ↗
    all of them →
  2. background

    Google touts #1 benchmark rankings for the new models — Google's blog post says Gemini 3.8 Flash TTS took the #1 overall spot on Hume AI's Voice Design Benchmark and led in accent modeling, while both new models topped Hume AI's Overall Quality Index and ranked highly in blind Voice Arena preference tests across languages including Japanese, Brazilian Portuguese, Vietnamese, MSA, Mexican Spanish and Hindi.

  3. 2

    Google launches Gemini 3.8 Flash TTS and Flash-Lite TTS

    Google and Google DeepMind announced two new text-to-speech models offering custom voice design, voice replication with consent verification, scene direction, and support for over 100 languages, available via Google AI Studio, the Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids.

    “Introducing Gemini 3.8 Flash and Flash-Lite TTS, our new SOTA text to speech model with: - a new voice design experience - 2,000+ production ready voices - voice replication - support for 100 languages - voice remixing (soon) - #1 spot on Hume AI's voice benchmarks and more!!”
    — @officiallogank
    1. first by HN Best, 11h ago · also HN Frontpage, Google DeepMind

      1 more headline
    2. first by Unite.AI, 11h ago

    1 more claim →

What people are saying 19 voices from 3 sites · best of 30 · verbatim