Live data from Hacker News

Gemini 3.8 Live and 3.8 Live Extended Thinking

blog.google

141–150 of 335 posts

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#141
post #85

Gemini is underrated in that it produces the only prose that is somewhat bearable to read.

I was surprised when (finally) trying out Claude how much I preferred Gemini's way of communicating. I wont argue Claude is better at coding, but for knowledge work, I had to dig through Claude output to find what I actually wanted. At times, it even felt borderline incomprehensible.

Opus specifically talks as if having a stroke. 4.6 was the last version that was pleasant to work with

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#142
post #9

I wonder when/if we’ll see Gemini beating Fable and Astra. Last year I would have confidently bet Google will overtake the others just because they have the data, the hardware (TPUs) and a fat advertising money pipe and yet they are still behind. Anyone anonymous at Google want to hint when Gemini 4 will be out?

Considering how far backwards Google has gone since the release of the 3.x models, the brain drain they've let happen, the lacklustre software ecosystem and their track record with product management ... I would say they're more likely to drop trying to compete at the frontier and try to focus on something else instead.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#143
post #99

Earlier quoted context omitted.

I actually did (was going to travel internationally), and it wasn't as useful as you'd think. I would be talking to someone, and in the background someone else would be talking, and it would translate both people. Only worked in a 1:1 in a quiet place. Still, can't complain for free.

"Only worked in a 1:1 in a quiet place" well then its not model problem

It's not a model problem, but it is a SW problem. It should be able to distinguish nearby people from people farther away and give me options to set a threshold on who to include.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#144

My first language is Afrikaans, which is a somewhat niche language and hard to find teachers/conversation buddies outside South Africa. (I live in USA now) I've been using Gemini to live chat in Afrikaans and do impromptu Afrikaans grammar lessons during my solo drives around town. It is phenomenal at speaking the language - like, it really shocks my family members when they hear it. This is probably the most joy I g…

I have a similar experience using Gemini for quick Catalan translations for iOS apps given enough context.

I once asked it to summarize The Hobbit in Catalan to explain it to my daughter before sleep. I was expecting a lot of mistakes as I see regularly if I ask anything in my native language when using GPT or Claude, but it was surprisingly good. I was going just to kind of skim ahead and retell it my own way, but ended up almost saying it verbatim because it was good already.

She loves Zelda so I asked it to explain the story of Breath of The Wild keeping the original names, and to make it fun, etc.. I was surprised again. I did retell some bits in my own style and taste but it is very convincing.

I haven't tried Catalan on newer models like GTP-6 Astra or Fable tho. We have all these benchmarks based on software development, and AGI, etc.. but it would be cool to have some language benchmarks for different communities.

As I work in english and use them in english, I wonder if using LLMs in a different language to code renders a different result as well. Like, if some of these benchmarks were made in other languages, would the result be similar.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#145

Gemini is underrated in that it produces the only prose that is somewhat bearable to read.

For heavyweight work I have been using Astra, but for rabbit holes and brain storming Gemini is far more enjoyable to interact with. I'm worried in their push to catch up on the SOTA front, it's going to lose that natural sounding touch it currently has.

How do you know you're using Astra?

My ChatGPT env only says "low", "medium", "high".

Is this a "pro" thing? I have totally no idea what I'm talking to, so actually I'm thinking of stopping my plan. Gemini and Claude are much more clear about it.

Anyway, I like the speed at which Gemini responds so indeed for simple things it is preferable.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#146
post #36

Earlier quoted context omitted.

Yeah its good. Reasonably priced too (at current prices, if they do raise them in January I would stop recommending it). 3.7/3.8 were good releases.

To me it's a good replacement for search engines. I ask it things like 'If redshifting destroys energy ala Noether, than how can we say that time is reversible or that entropy will find an equilibrium?' and it will not only explain, but make nice interactive diagram/toys to help. A regular search engine would have taken me hours to find an answer. However, if I want it to DO something then Gemini is in absolute last…

I’ve mapped my iPhone Action Button directly in to Google App AI Mode (which is Gemini 3.8 but faster inference than the Gemini App, due to harness)

It’s for this exact kind of scenario where a random question pops in to my head.

Plus, it’s the most grounded by real live data of all the chatbots.

Silicon Valley people are majorly sleeping on Google Search AI Mode.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#147

Earlier quoted context omitted.

For coding models? I don't think Google is motivated to fight in that market. There's no incentive for them. Ask yourself, how does Google -- a company that famously does everything -- benefit from SWEs outside of Google having access to powerful coding models? They would just be competition. Famously, Google just eventually discards almost all businesses that don't have the same fire hose of revenue that ads does. S…

> Ask yourself, how does Google -- a company that famously does everything -- benefit from SWEs outside of Google having access to powerful coding models? Catch is, even Googlers internally do not have access to top-tier models. (or did not until recently, when apparently Claude was made accessible to the SWEs internally).

Googlers now have internal access to frontier models.

Funny enough, after trying it, I went back to G3.8

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#148

Earlier quoted context omitted.

Mostly because it answers quickly and is more agreeable (too agreeable at times). Meanwhile Claude and Astra like to couch all their agreements with caveats and provisos.

“caveats and provisos” makes me think of Robin Williams’ genie imitating William F. Buckley Jr.

This is accurate.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#149
post #9

I wonder when/if we’ll see Gemini beating Fable and Astra. Last year I would have confidently bet Google will overtake the others just because they have the data, the hardware (TPUs) and a fat advertising money pipe and yet they are still behind. Anyone anonymous at Google want to hint when Gemini 4 will be out?

For me 3.8 has been good enough that I don’t think I’ll be extending my Claude subscription.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#150

They should just give up at this point, it's just embarrassing to watch. As PrimeTime said; these are the guys that invented the 'T' in 'GPT', that deployed their first TPU in 2015, that is using billions on AI - and they are beaten by 300 people startup named Moonshot AI even. People are going to write books about this complete fumble.

Yeah, you should definitely sell all Google stock you may hold, immediately, to me, before it's too late. I'm just a sucker, I'm happy to buy from you.
Post reply on HN