Google AI

@GoogleAI

Introducing our most advanced Gemini Audio models yet 🗣 Gemini 3.8 Live and 3.8 Live Extended Thinking let you speak, collaborate, and execute tasks seamlessly, meaning conversing with AI just got a lot more natural. So, what’s the difference between these two models? Let’s break it down: — Gemini 3.8 Live is built for scale, speed, and cost efficiency. It can handle mid-sentence interruptions, transitions across 97 languages on the fly, and understands visual context. Figure out how to fix a broken bike chain, or deal with a leaky pipe just by pointing your camera at the problem area in Search Live for step-by-step audio instructions. — Gemini 3.8 Live Extended Thinking goes one step further to bring increased intelligence to your most complex tasks. It reasons and speaks in parallel, even narrating its progress as it works. This lets it handle multi-step, behind-the-scenes projects, like planning an event, without ever losing the conversational flow. Watch how Gemini 3.8 Live combines real-time video and voice inputs in Search Live to tackle hands-on DIY plumbing tasks step by step 👇
打开原帖#511482
  1. Frontier

    Nathan Lambert: An basic idea in scaling RL: Can we allocate more compute to the hard…
  2. Frontier

    Nathan Lambert: Excited this paper is out (and surprised I haven't seen someone pursu…
  3. Frontier

    Matt Shumer: Guys, these insane setups are fun and all, but they don't actually ma…