Artificial Analysis publishes independent speed and quality benchmarks
1 Today 11:50 AM · 11h ago · 1 article · 7 posts · 2 sources · development 1 of 2
Third-party AI benchmarking tracker Artificial Analysis reported Gemini 3.8 Flash TTS processes 44.1 characters per second (about 2.7x realtime) versus 40.2 for Flash-Lite, placing Google's models behind faster competitors like Falcon 2, while noting the Flash model debuts at #1 on its Pronunciation Robustness Benchmark and #2 on its Provider Voice Arena Leaderboard.
“Gemini 3.8 Flash TTS processes 44.1 characters per second, compared to 40.2 characters per second for Gemini 3.8 Flash-Lite TTS, approximately 2.7x and 2.4x faster than realtime, respectively.”
@artificialanlysGoogle Developer of Gemini modelsGoogle DeepMind AI research lab behind the models@officiallogank Google AI product figureArtificial Analysis Independent AI benchmarking tracker
The whole story articlesposts the bright band is this development · numbered dots are the others · click one to jump
What was reported 1 claim about this development
-
first by The Next Web, 11h ago · also Google, Simon Willison, Simon Willison's Weblog, The Decoder, MarkTechPost
4 more headlines
- Google releases Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, its “most expressive audio generation models yet”, with support for more than 100 languages Google · 11h ago
- Gemini 3.8 TTS Playground Simon Willison · 10h ago
- Google's new Flash TTS models let you design AI voices from scratch using text descriptions The Decoder · 9h ago
- Google Releases Gemini 3.8 Flash TTS and Flash-Lite TTS With Prompt-Based Voice Design marktechpost.com · 7h ago
What people said 24 voices · best of 30 · verbatim
-
The new Gemini 3.8 TTS models are super-cheap and can generate conversations between multiple voices (from 2,000+, or you can clone your own) - I built a little UI for it, then had Claude knock up a script where two pelicans debate moving to Pacifica Pier
-
Here's a video of my Emotive Audiobook Creator, KeenLore, a locally hosted web app:https://www.youtube.com/watch?v=WAeHgE94rVoNo cloud, no tokens to pay. Reads a book using a full cast of characters. Quotation attribution detection (for my novel) is at 97.2% accuracy (485/499 quotes identified and assigned correctly). The autofill of character…
-
H
Gemini 3.8 text-to-speech Link: https:// blog.google/innovation-and-ai/ models-and-research/gemini-models/gemini-3-8-text-to-speech/ Discussion: https:// news.ycombinator.com/item?id=4 9817615
-
Forget the leaderboard wins. The real headline in Google's Gemini 3.8 TTS launch is that cloning a voice from a sample lasting 30 seconds is now a standard developer feature. Google's safeguard is a matching spoken consent recording from the voice owner, plus SynthID watermarks and C2PA credential...
-
I vibe coded a playground UI for trying this out. The conversation mode is neat, and it's very expensive - most of my experiments have cost less than a cent.
-
N
Gemini 3.8 text-to-speech: https:// blog.google/innovation-and-ai/ models-and-research/gemini-models/gemini-3-8-text-to-speech/ Discussion: http:// news.ycombinator.com/item?id=4 9817615
-
We're launching Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS ⚡️ Our most expressive audio models yet let you create custom voices across 100+ languages or pick from 2,000+ ready-to-use ones. You can direct back-and-forth conversations, guide the delivery line-by-line, and add natural cues li...
-
I direct my own extended daydream Star Trek fanfic (okay, I'm on season 2 episode 17) and recently I looked to see if I could have each scene file be read aloud a la an audiobook or radio drama.Getting GPT-Live to have unique enough voices and to be expressive with how I imagine the voices going in my head is hard to direct, there's not enough…
-
Google has released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS. Gemini 3.8 Flash TTS debuts at #1 on our Pronunciation Robustness Benchmark and #2 on our Provider Voice Arena Leaderboard Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS are @GoogleDeepMind's latest Text to Speech models, ...
-
It's giving me an error when I try to generate a voice with Voice Design in AI Studio. It also says voice replication isn't available in my region.Also weird that there are no "neutral gender" voices in the English language. There's also limited "use cases," like the "Gaming" use case is empty?And there's no pricing listed anywhere.I don't know, I…
-
Gemini 3.8 Flash TTS processes 44.1 characters per second, compared to 40.2 characters per second for Gemini 3.8 Flash-Lite TTS, approximately 2.7x and 2.4x faster than realtime, respectively. Both models remain behind faster Text to Speech models we track, such as Falcon 2 at 204.9 characters per...
-
Pet peeve on Google's AI rollouts: there's no alignment across the three platforms they have, consumer, prosumer, cloud. Scroll to the end of every release, including this one, and you'll see different availabilities. The fun part is the models don't even have the same capabilities across platforms! Omni Flash, last I tried and read the docs, is…
-
I was invited to test Gemini 3.8 Flash TTS. This is currently my favorite model. The voice outputs from this model are extraordinary. I never really rated AI voice generation because of the uncanniness associated with low quality. This Gemini IMO crosses that chasm, like Opus 4.5 did with coding,...
-
> Voice replication: Recreate consistent vocal profiles from just a 30-second audio sample of your voice or a voice you have the rights to use, backed by built-in consent verification, SynthID watermarking, and C2PA credentials to protect both developers and their vocal talent.I guess voice cloning is widely enough available now from other…
-
introducing Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, our most expressive audio generation models yet these models enable creators, developers, and enterprises to create richer, more expressive audio experiences try them via the Gemini API and in AI Studio: https://aistudio.google.com/ ...
-
An education account I have lists 3.1 Pro, and 3.6 Thinking and Flash as the available models in the app.My personal account, 3.6 Flash Lite and 3.5 Thinking.Meanwhile, I can go hog wild and drain my bank account on GCP. I don’t though, because my family has to eat.
-
Create and deploy custom audio with our new text-to-speech models: 🔵 Gemini 3.8 Flash TTS: Design unique voices with distinct accents and characteristics. 🔵 Gemini 3.8 Flash-Lite TTS: Built for efficiency and scale, choose from your created styles or our expansive production-ready library.
-
Google is spreading too thin, as gemini isn't really that intelligent.They are creating gemini SOTA (not really any more), flash versions, text-to-speech, video (omni), etc.I can see they want to create an ecosystem, but I see no focus in any one area.
-
Introducing Gemini 3.8 Flash and Flash-Lite TTS, our new SOTA text to speech model with: - a new voice design experience - 2,000+ production ready voices - voice replication - support for 100 languages - voice remixing (soon) - #1 spot on Hume AI's voice benchmarks and more!!
-
My primary use case for TTS is converting written content (blogs, articles, etc.) in to clips I can listen to on the go.Is there a good browser extension that does this with a flexible TTS backend? I know Qwen, Kokoro, and VibeVoice all have decent quality..
-
Rolling out starting today: — Developers: Both models in @GoogleAIStudio and the Gemini API — Consumers: Gemini 3.8 Flash TTS in @Gemini_Notebook and Gemini 3.8 Flash-Lite TTS in Google Vids — Coming soon: Both models in Gemini Enterprise
-
Related, for embedding small models, this lib is incredible.Having a voice under 1Mo is crazy, even if it sounds robotic.
-
Can you hear that? Our Gemini Audio family is getting louder 🔊 We're introducing two of our most expressive audio generation models yet from @GoogleDeepMind: Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS.
-
Interesting to see that they are publishing a new tts model. I remember that they refused to release one of them a few years ago because they were too afraid of abusing it. Now they just release it without any much thought lol
All 2 developments of Google launches Gemini 3.8 Flash TTS with custom voice… →
NewswiresYouTubeHacker NewsXMastodonGoogle News