conv.

All stories
AIQuiet 8d · day 11

Google launches Gemini 3.8 Live voice models, undercuts OpenAI on price

New audio-to-audio models top speech benchmarks and reason in the background while still talking, at a fraction of OpenAI's cost.

What to know

  • Gemini 3.8 Live and Extended Thinking are Google's newest real-time voice AI models, live now via the Gemini API, Workspace, and the Gemini app.
  • Extended Thinking ranks #1 overall on Artificial Analysis' Speech to Speech Quality Index (82.6) and leads agentic voice benchmarks including τ-Voice and Big Bench Audio.
  • At $1.38 per hour of conversation, the models undercut OpenAI's GPT-Live-1 on price, though OpenAI's full-duplex model may still sound more natural.
  • Models support 97 languages with mid-conversation switching, near real-time visual context, and background/async tool execution during live conversation.

“Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet.”

Google DeepMind, Company blog post · Google DeepMind blog ↗ · Sep 14

Google DeepMind Developer of Gemini 3.8 Live modelsLogan Kilpatrick Google AI product leadOpenAI Competitor, maker of GPT-Live-1Artificial Analysis Independent AI benchmarking organization

Google launches Gemini 3.8 Live voice models, undercuts OpenAI on price
blog.google

How it unfolded 3 developments, newest first · click a bar or a number to jump articlespostscomments

Peak 57 pieces in 3h at Sep 15, 12 PM; 110 pieces over 11 days (41 articles · 30 posts · 39 comments) Sep 15, 12 PM — 57 pieces · 22 articles · 24 posts · 11 comments — Newswires 20, X 20, Hacker News 13, +2 moreSep 15, 3 PM — 18 pieces · 3 articles · 1 post · 14 comments — Hacker News 14, Google News 2, Mastodon 1, +1 moreSep 15, 6 PM — 7 pieces · 3 articles · 1 post · 3 comments — Hacker News 4, Google News 3Sep 15, 9 PM — 3 pieces · 3 comments — Hacker News 3Sep 16, 12 AM — 3 pieces · 1 article · 2 comments — Hacker News 2, Google News 1Sep 16, 3 AM — 3 pieces · 1 post · 2 comments — Hacker News 2, Mastodon 1Sep 16, 6 AM — 3 pieces · 1 post · 2 comments — Hacker News 2, Mastodon 1Sep 16, 9 AM — 2 pieces · 2 comments — Hacker News 2Sep 16, 12 PM — quietSep 16, 3 PM — 2 pieces · 2 articles — Google News 2Sep 16, 6 PM — 4 pieces · 4 articles — Google News 4Sep 16, 9 PM — quietSep 17, 12 AM — quietSep 17, 3 AM — 4 pieces · 2 articles · 2 posts — Mastodon 2, Google News 1, Newswires 1Sep 17, 6 AM — quietSep 17, 9 AM — 3 pieces · 3 articles — Google News 3Sep 17, 12 PM — quietSep 17, 3 PM — quietSep 17, 6 PM — quietSep 17, 9 PM — quietSep 18, 12 AM — quietSep 18, 3 AM — quietSep 18, 6 AM — quietSep 18, 9 AM — quietSep 18, 12 PM — 1 piece · 1 article — Google News 1Sep 18, 3 PM — quietSep 18, 6 PM — quietSep 18, 9 PM — quietSep 19, 12 AM — quietSep 19, 3 AM — quietSep 19, 6 AM — quietSep 19, 9 AM — quietSep 19, 12 PM — quietSep 19, 3 PM — quietSep 19, 6 PM — quietSep 19, 9 PM — quietSep 20, 12 AM — quietSep 20, 3 AM — quietSep 20, 6 AM — quietSep 20, 9 AM — quietSep 20, 12 PM — quietSep 20, 3 PM — quietSep 20, 6 PM — quietSep 20, 9 PM — quietSep 21, 12 AM — quietSep 21, 3 AM — quietSep 21, 6 AM — quietSep 21, 9 AM — quietSep 21, 12 PM — quietSep 21, 3 PM — quietSep 21, 6 PM — quietSep 21, 9 PM — quietSep 22, 12 AM — quietSep 22, 3 AM — quietSep 22, 6 AM — quietSep 22, 9 AM — quietSep 22, 12 PM — quietSep 22, 3 PM — quietSep 22, 6 PM — quietSep 22, 9 PM — quietSep 23, 12 AM — quietSep 23, 3 AM — quietSep 23, 6 AM — quietSep 23, 9 AM — quietSep 23, 12 PM — quietSep 23, 3 PM — quietSep 23, 6 PM — quietSep 23, 9 PM — quietSep 24, 12 AM — quietSep 24, 3 AM — quietSep 24, 6 AM — quietSep 24, 9 AM — quietSep 24, 12 PM — quietSep 24, 3 PM — quietSep 24, 6 PM — quietSep 24, 9 PM — quietYesterday, 12 AM — quietYesterday, 3 AM — quietYesterday, 6 AM — quietYesterday, 9 AM — quietYesterday, 12 PM — quietYesterday, 3 PM — quietYesterday, 6 PM — quietYesterday, 9 PM — quietToday, 12 AM — quietToday, 3 AM — quietToday, 6 AM — quietToday, 9 AM — quietToday, 12 PM — quietToday, 3 PM — quiet 1–3
Sep 16Sep 17Sep 18Sep 19Sep 20Sep 21Sep 22Sep 23Sep 24yesterdaynow · 5:59 PM ET
  1. 3

    The Decoder frames launch as a price attack on OpenAI's GPT-Live-1

    The Decoder reported that the new models top the Artificial Analysis speech-to-speech leaderboard and cost $1.38 per hour of voice conversation, significantly undercutting OpenAI's GPT-Live-1, while noting OpenAI's model may still sound more natural due to full duplex audio.

    “Say hello (literally) to Gemini 3.8 Live and 3.8 Live Extended Thinking, our new SOTA live audio models, available with frontier price + performance. 3.8 Live supports 97 languages (can seamlessly switch), async tool calls, and more!”
    — @officiallogank, Google AI product lead · source
    1. first by Rohan's Bytes, 11d ago · also TestingCatalog AI News, Neowin, Thurrott, RuntimeWire, Google, AI Weekly +5

      12 more headlines
    2. first by SiliconANGLE, 11d ago · also Search Engine Land, The Decoder, Simon Willison, Simon Willison's Weblog, Search Engine Roundtable, Memeburn

      6 more headlines
    • Gemini Notebook’s latest update makes it a better study companion "Google is adding voice conversations, lecture recording, interactive quizzes, and short video overviews to its AI-powered notebook. " by Pranob Mehrotra / via Digital Trends # AI # tech # NotebookLM # GeminiNotebook # productivity # learning # knowledgeManagement # Google # study…

      flaberenne@mastodon.socialMastodon11d ago1▲view on Mastodon ↗
    2 more of the top 3 · 61 posts in this stretch
    • I have a similar experience using Gemini for quick Catalan translations for iOS apps given enough context.I once asked it to summarize The Hobbit in Catalan to explain it to my daughter before sleep. I was expecting a lot of mistakes as I see regularly if I ask anything in my native language when using GPT or Claude, but it was surprisingly good…

      LluisGerardHacker News10d agoview on Hacker News ↗
    • Google has released Gemini 3.8 Live, its new Speech to Speech model, with the Extended Thinking (High) variant debuting at #1 on the Artificial Analysis Speech to Speech Index at 82.6, and #1 on our Tau Voice benchmark implementation at 68.6% Gemini 3.8 Live is @GoogleDeepMind's successor to Gemin...

      @artificialanlysX11d agoview on X ↗
    all of them →
  2. 2

    Google details background reasoning for developers

    Google's API documentation describes Extended Thinking as a high-reasoning audio-to-audio model that processes background reasoning and asynchronous tool calls while streaming continuous audio, requiring developers to update client state management to handle async reasoning signals.

  3. 1

    Google launches Gemini 3.8 Live and Extended Thinking

    Google DeepMind introduced two new live dialogue models offering near real-time reasoning, visual context processing, mid-conversation language switching across 97 languages, and background tool execution; available immediately via the Gemini API, Workspace, and the Gemini app.

    1. first by Breakingthenews.net, 11d ago · also Seeking Alpha, PYMNTS, Deccan Herald, TechRepublic, Inshorts

      5 more headlines
    2. 2 outlets first by HN Best, 11d ago · also HN Frontpage · read ↗

Also covered reported alongside — the timeline has no entry for these yet

  1. first by Mashable India, 9d ago · also NewsBytes

    1 more headline

and 2 smaller pieces

What people are saying 21 voices from 3 sites · best of 61 · verbatim