Gemini 3.8 Live and 3.8 Live Extended Thinking

(blog.google)

97 points | by leumon 2 hours ago ago

49 comments

  • Zsfe510asG an hour ago

    Gemini is underrated in that it produces the only prose that is somewhat bearable to read.

    • WarmWash 41 minutes ago

      For heavyweight work I have been using Astra, but for rabbit holes and brain storming Gemini is far more enjoyable to interact with.

      I'm worried in their push to catch up on the SOTA front, it's going to lose that natural sounding touch it currently has.

      • alansaber a few seconds ago

        Agreed. My impression is that the more verbose output of sol, astra etc is that it helps it steer itself on long running tasks (but is worse for the human user to read)

      • drivebyhooting 4 minutes ago

        Mostly because it answers quickly and is more agreeable (too agreeable at times).

        Meanwhile Claude and Astra like to couch all their agreements with caveats and provisos.

      • martythemaniak 16 minutes ago

        Good news then, I don't think they're in a hurry to catch up to SOTA.

    • ympb121 10 minutes ago

      Wondering if people have managed to have Gemini in-front of other models like claude/codex models and only interact with that. Having Gemini act as a pure human/llm translator.

    • throwaw12 8 minutes ago

      Does anyone know if there is a dedicated model which makes Claude output nore human readable and less slop?

      Lately it became load-bearingly-reality-difficult to not only read, but to comprehend the Claude output

  • deviation an hour ago

    Not a great impression to have your demo video demonstrate how one of your 'most advanced' AI models loses to the most common check-mate pattern in all of chess.

    • monroewalker 40 minutes ago

      Seems more than good enough for a live model though! I can imagine this demo being extended to be a lot nicer to play with. You can just feed the model engine analysis and it can make as high of quality moves as needed. No longer any correlation between the model's understanding of the position and the moves that would be made but I think that's still a really nice improvement when thinking about this as adding live voice interaction to existing chess vs computer functionality rather than adding chess to possible interactions with the latest live voice model.

  • doodlesdev an hour ago

    Gemini's Live Mode is already much better than GPT Voice in my personal experience, even though it was much dumber. It really does feel like talking to a real person. ChatGPT keeps humming to whatever I say and has some weird voices.

    Excited to try this out! Shame on Google for not releasing Gemini 3.8 for Google AI Plus users yet, though.

    • ilaksh 2 minutes ago

      OpenAI just released the new full duplex mode to the API as gpt-live-1 or something like that. Very realistic.

  • sahaskatta an hour ago

    Our company's Google Workspace Business only offers 3.6 flash & thinking in the Gemini App. Has anyone else seen 3.7 or 3.8 roll out?

    • cnobody 40 minutes ago

      I have access to both, benchmarks are actually better on 3.7 for my task, but happy improvement over the others.

  • rdtsc 44 minutes ago

    I wonder when/if we’ll see Gemini beating Fable and Astra. Last year I would have confidently bet Google will overtake the others just because they have the data, the hardware (TPUs) and a fat advertising money pipe and yet they are still behind. Anyone anonymous at Google want to hint when Gemini 4 will be out?

    • WarmWash 35 minutes ago

      As an "everyday mans AI" I'd say 3.8 Flash definitely already has. Smart enough for the vast swath of people, and only slightly eeked out by Astra(Max) on vision capabilities, like the kind of "Point your camera at something and ask questions" that non-tech people like to do. It's crazy fast and very compute light, so not getting bogged down constantly.

      I can't think of a better general purpose model than 3.8 flash right now. It also writes more naturally than the other big models too.

      • mchusma 16 minutes ago

        Yeah its good. Reasonably priced too (at current prices, if they do raise them in January I would stop recommending it). 3.7/3.8 were good releases.

    • _s_a_m_ 42 minutes ago

      If Google didnt have their ad buisness theiy'd be out by now. They're like BlackBerry and Nokia at this point almost.

      • nolok 31 minutes ago

        Not sure what your comment mean in the context of parent's comment.

        As opposed to what, them not having it and burning money that isn't their instead like openai and anthropic? At least Google is feeding itself instead of having to create a bubble to stay alive

        • verdverm 27 minutes ago

          Google spent most of their cash, they are now taking loans for data centers too. They recorded their first quarter of negative cash flows ever

          • runako 8 minutes ago

            Google has in excess of $121 billion of cash (& equivalents), net of total debt.

            Big rich companies take on debt for reasons that are sometimes inscrutable from the outside. Recently, they have been borrowing for ~5%, about a half point above what the US government gets for 10-year Treasuries.

            Apple has been financing operations with debt for a number of years as part of a complex optimization plan.

            No, Google is not broke.

          • nolok 8 minutes ago

            Which is still a better position that the others? My point is you can't consider that a bad thing if you think it's ok for their competitors in the field. And if you don't and your judge them equally, then at least Google has its own cash glow and could turn the gas off at any point to go back to printing money while they have not choice.

      • haberdasher 33 minutes ago

        Maps, Waymo, TPUs, YouTube, Docs, GMail, Cloud, Android, Chrome, Photos...

  • samuelknight 28 minutes ago

    I have been looking for a model that's good for GUI testing. Original computer use isn't right because it's a slow screenshot loop, which doesn't capture transition and animation. Docs says this one does up to 1 FPS. That might be fast enough. If not now, we must be within a few months of high enough sample rates to do it.

  • 740273730191 30 minutes ago

    They can't even vibecode a working VS Code extension for Gemini.

    Nothing but constant errors with cryptic messages.

    • verdverm 24 minutes ago

      It's a slop factory over there apparently. We were sent the greatest slop deck of all time from their sales team. We now have a :cursed-claude: from a slide where they said "we have access to state of the art models like Claude 3" and nanobanana's interpretation of what Claude looks like as a person. It was clear the person had only read a handful of the nearly 40 slides

  • blovescoffee 36 minutes ago

    Great tech but the voice is like nails on a chalkboard to me

    • tantalor 25 minutes ago

      Which one? I like Eclipse

  • smithcoin 26 minutes ago

    Did anybody watch the Primeagen's video on Google bag-fumbling? Interesting they released on the same day!

  • attels33 an hour ago

    When will it be available on Vertex?

    • verdverm 22 minutes ago

      I've been asking the same about the open weight models, we're buying our tokens from others now, though I think those people are renting hardware from Google in the end anyway

    • SomeonesAccount an hour ago

      vertex is dead, for good reason too

      • bilarikan 6 minutes ago

        Would you mind expanding on this? I thought Google changed the name to 'Gemini Enterprise Agent Platform', and altered focus to 'agent governance' workflows, but that there were no breaking changes from what was offered with Vertex AI.

  • mvdtnz 29 minutes ago

    My Gemini app is still stuck at 3.5 Flash-lite and 3.6 Flash so I truly don't understand how Google rolls this stuff out. I don't use Gemini for anything serious so I'm not going to use the API, but it's my go-to for just searching basic information (replacing google search) because it's so darn fast.

  • lostmsu 30 minutes ago

    So I am building a voice assistant to control AI harnesses, and recently tried switching from GLM 5.3 Flash to Gemini 3.8 Flash because of higher tok/s and better rate limits. Before that I also used Kimi K3 and DeepSeek-V4-Flash-0731.

    Let me tell you unlike every other mentioned model Gemini 3.8 Flash trial had to be reverted the same day. Instead of simply delegating tasks it would invent additional requirements and implementation details it knew nothing about and no amount of convincing not to do it would work. That's the first time a model failed on me so spectacularly despite having practically same Artificial Analysis Intelligence Index as another model that just worked (and higher than working DS Flash).

    The reason I think it is relevant is: Live is likely even stupider model in every way possible (except hearing better than separate STT). So beware using it for agentic scenarios.

  • glimshe 33 minutes ago

    I'm disappointed with "Extended Thinking" for 3.8 Flash. On the plus side, it's a strong general-purpose model and the cost-benefit is still compelling.

    However, the "Extended Thinking" should be renamed to "Slightly Extended Thinking". Considering that it's the maximum thinking option for Gemini Flash in the chat UI, it doesn't actually think a whole lot, leading to an uncomfortably high number of incorrect/poor replies.

  • tiahura 39 minutes ago

    Ensure transparency with SynthID watermarking

    All audio generated by our AI products is watermarked with SynthID. This imperceptible watermark is woven directly into the audio output, ensuring AI-generated content remains detectable to help prevent misinformation. For details on our approach to safety and responsibility, review the model card.

    • hajile 30 minutes ago

      It’s little to do with misinformation and much to do with trying to keep their model from collapsing from ingesting too much slop.

  • varispeed an hour ago

    "Thinking" Good one.

  • bronlund 34 minutes ago

    They should just give up at this point, it's just embarrassing to watch.

    As PrimeTime said; these are the guys that invented the 'T' in 'GPT', that deployed their first TPU in 2015, that is using billions on AI - and they are beaten by 300 people startup named Moonshot AI even. People are going to write books about this complete fumble.

    • password54321 29 minutes ago

      None of the startups are profitable. What exactly are they getting beaten at?

      My advice is to listen less to brainrot 'influencers' that optimise for engagement through sensationalism.

      • bronlund 25 minutes ago

        Intelligence.

        They have "unlimited" resources and has researched AI since the very beginning - PageRank is a form of AI even. And still, Gemini is behind Claude, GPT, Grok, Muse, GLM, Kimi and is maybe on par with DeepSeek?

        As I said, it is embarrassing.

        • password54321 7 minutes ago

          If RSI is achievable, it will leapfrog everything produced so far and so it will make sense to focus on RSI instead of incremental improvements for your top model. Startups need investment and need to show progress. Google does not at the moment need to take lead in the current race.

        • Forgeties79 13 minutes ago

          > Grok

          No one is behind grok. It literally has "be funny and irreverent when appropriate" (whatever the hell "when appropriate" means for them) baked into the system prompt. To me, that is all you need to know about how useful it is.

          No serious people use it and the numbers bear it out tbh. It has the smallest market share of the "big companies" for a reason - and it's by a very, very large margin (~2.5% last I checked).

    • mattlondon 12 minutes ago

      Have you used Google search at all on the past few months? Every single search brings up a live chat prompt. They're serving fast AI to billions of users at huge scale everyday

      And they're making money doing it.

      Perhaps they don't have the best coding model right now (although 3.8 flash is arguably SOTA at some benchmarks), but is that the be-all and end-all of AI? Only coding matters?

      • thereitgoes456 5 minutes ago

        It’s so annoying that everyone just points to the Artificial Analysis index (or even worse, Epoch AI, where part of the score is how good the AI is at chess) as a proxy for “how good” the model is.

    • mchusma 15 minutes ago

      Their live models have been and continue to be at the frontier. I like them a lot!

    • dude250711 21 minutes ago

      It makes Meta look not that bad.

    • fileeditview 31 minutes ago

      And yet they might become one of the winners "in the end" because they have near infinite money and others have not. I will drink tea and watch the show.

      • tonfa 20 minutes ago

        > they have near infinite money and others have not

        Given the very high margins on inference, once volume is large enough the other can also start printing enough money.