logoalt Hacker News

Gemini 3.8 Live and 3.8 Live Extended Thinking

167 pointsby leumontoday at 5:38 PM107 commentsview on HN

Comments

Zsfe510asGtoday at 6:46 PM

Gemini is underrated in that it produces the only prose that is somewhat bearable to read.

show 1 reply
Havoctoday at 7:54 PM

Just gave it a try - very solid release.

Copes well with thick accent, voices are pleasant and latency seems low.

Oh and I can actually use it on a workspace account - which for most of the recent releases was an account stuck in limbo. Not personal enough for personal offering, not enterprise enough for enterprise.

Well done G - will definitely be using this

show 1 reply
rdtsctoday at 6:52 PM

I wonder when/if we’ll see Gemini beating Fable and Astra. Last year I would have confidently bet Google will overtake the others just because they have the data, the hardware (TPUs) and a fat advertising money pipe and yet they are still behind. Anyone anonymous at Google want to hint when Gemini 4 will be out?

show 2 replies
sahaskattatoday at 6:47 PM

Our company's Google Workspace Business only offers 3.6 flash & thinking in the Gemini App. Has anyone else seen 3.7 or 3.8 roll out?

show 1 reply
deviationtoday at 6:43 PM

Not a great impression to have your demo video demonstrate how one of your 'most advanced' AI models loses to the most common check-mate pattern in all of chess.

show 1 reply
doodlesdevtoday at 6:48 PM

Gemini's Live Mode is already much better than GPT Voice in my personal experience, even though it was much dumber. It really does feel like talking to a real person. ChatGPT keeps humming to whatever I say and has some weird voices.

Excited to try this out! Shame on Google for not releasing Gemini 3.8 for Google AI Plus users yet, though.

giancarlostorotoday at 8:12 PM

I'm wondering if Google intends to drop the next major version of Gemini Pro as a total bombshell drop to make Anthropic and OpenAI panic. They seem to be taking their sweet time on frontier model updates.

stranded22today at 8:15 PM

Google need to allow saving history and exclude it as training data. I will not use it seriously until this is resolved.

show 1 reply
samuelknighttoday at 7:07 PM

I have been looking for a model that's good for GUI testing. Original computer use isn't right because it's a slow screenshot loop, which doesn't capture transition and animation. Docs says this one does up to 1 FPS. That might be fast enough. If not now, we must be within a few months of high enough sample rates to do it.

740273730191today at 7:05 PM

They can't even vibecode a working VS Code extension for Gemini.

Nothing but constant errors with cryptic messages.

smithcointoday at 7:09 PM

Did anybody watch the Primeagen's video on Google bag-fumbling? Interesting they released on the same day!

show 1 reply
ghoshbishakhtoday at 7:45 PM

I love talking to chatgpt voice mode. Voice to voice AI is the only big leap that I see after the RL trained coding models.

show 1 reply
blovescoffeetoday at 6:59 PM

Great tech but the voice is like nails on a chalkboard to me

attels33today at 6:49 PM

When will it be available on Vertex?

show 1 reply
glimshetoday at 7:02 PM

I'm disappointed with "Extended Thinking" for 3.8 Flash. On the plus side, it's a strong general-purpose model and the cost-benefit is still compelling.

However, the "Extended Thinking" should be renamed to "Slightly Extended Thinking". Considering that it's the maximum thinking option for Gemini Flash in the chat UI, it doesn't actually think a whole lot, leading to an uncomfortably high number of incorrect/poor replies.

mvdtnztoday at 7:07 PM

My Gemini app is still stuck at 3.5 Flash-lite and 3.6 Flash so I truly don't understand how Google rolls this stuff out. I don't use Gemini for anything serious so I'm not going to use the API, but it's my go-to for just searching basic information (replacing google search) because it's so darn fast.

lostmsutoday at 7:06 PM

[delayed]

tiahuratoday at 6:57 PM

Ensure transparency with SynthID watermarking

All audio generated by our AI products is watermarked with SynthID. This imperceptible watermark is woven directly into the audio output, ensuring AI-generated content remains detectable to help prevent misinformation. For details on our approach to safety and responsibility, review the model card.

show 1 reply
varispeedtoday at 6:31 PM

"Thinking" Good one.

bronlundtoday at 7:02 PM

They should just give up at this point, it's just embarrassing to watch.

As PrimeTime said; these are the guys that invented the 'T' in 'GPT', that deployed their first TPU in 2015, that is using billions on AI - and they are beaten by 300 people startup named Moonshot AI even. People are going to write books about this complete fumble.

show 2 replies
chrystianpltoday at 8:54 PM

It is completely broken for me. After I ask a single question, it starts replying to itself in an infinite loop. It answers my question, then generates another reply to its own response, and keeps going. At some point, it even starts switching languages randomly.

show 1 reply