Live data from Hacker News

Gemini 3.8 Live and 3.8 Live Extended Thinking

blog.google

191–200 of 334 posts

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#192
post #9

I wonder when/if we’ll see Gemini beating Fable and Astra. Last year I would have confidently bet Google will overtake the others just because they have the data, the hardware (TPUs) and a fat advertising money pipe and yet they are still behind. Anyone anonymous at Google want to hint when Gemini 4 will be out?

I’m wondering if they even see a coding agent as a valuable prize. It’s a competitive market in a race to the bottom economically, hard to establish consistent differentiation and virtually zero switching cost for customers. I think they’ve made a shrewd move in focusing on search integration and everyday users (Gemini app) vs software power users. They have their corner and nobody is really competing with them, plus…

OpenAI/Anthropic are being valued at 25/50% of Alphabet respectively in secondary markets.

Maybe everybody is wrong about how valuable these companies will be, but atm it looks very dumb to not be competitive in coding.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#193

Earlier quoted context omitted.

Intelligence. They have "unlimited" resources and has researched AI since the very beginning - PageRank is a form of AI even. And still, Gemini is behind Claude, GPT, Grok, Muse, GLM, Kimi and is maybe on par with DeepSeek? As I said, it is embarrassing.

If RSI is achievable, it will leapfrog everything produced so far and so it will make sense to focus on RSI instead of incremental improvements for your top model. Startups need investment and need to show progress. Google does not at the moment need to take lead in the current race.

And yet they are burning through their famous cash on hand and taking on debt for data centers like everyone else.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#194

My first language is Afrikaans, which is a somewhat niche language and hard to find teachers/conversation buddies outside South Africa. (I live in USA now) I've been using Gemini to live chat in Afrikaans and do impromptu Afrikaans grammar lessons during my solo drives around town. It is phenomenal at speaking the language - like, it really shocks my family members when they hear it. This is probably the most joy I g…

I love Gemini in Google Maps for long drives. I start getting bored of music and podcasts and start grilling it with random questions I've always wondered about.

I do the same for improving my English skills. It's amazing!

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#195
i've been using gemini (api & pro) since last november last year. for creative writing compared to other llm, it's the best in capturing local nuance, it can even create jokes in my languages. But that just it, i cant rely on other work, hallucinate too often, the deep research are not reliable at all. Too many discussion i've had that it grasp main concept consistent but the supporting concept just plain hallucinate and not consistent. it's tested between pro & flash. This doesnt happen often on open weigh

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#196

Earlier quoted context omitted.

I’m curious how you guys keep track of each model’s coding capabilities. The landscape keeps changing. I don’t suppose you benchmark all frontier models every other month, right?

I use them. Daily. Gemini hasn’t been a contender by comparison for a long time.

I had a typical $20 Gemini plan that I just downgraded to their $5 plan (to keep access to some of the models). It had been so long since I let Gemini work on (or review) any code / design / html (anything) that I couldn't justify bothering to keep wasting money on it. It fell behind badly over the past year. Astra might as well be an alien super intelligence at code compared to Gemini. I enjoy talking to Gemini, it is very good at conversation, I get solid answers to everyday questions. I intend to keep the $5 plan indefinitely for basic use. I don't expect they'll ever resurface as a competitor in coding with Astra & Fable et al.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#197
post #162

Earlier quoted context omitted.

When comparing OpenAI and Claude thats pretty much true, but not Gemini... And have you tried Antigravity? Yikes

The CLI version of agy is great. Have you tried it?

Compared to gemini-cli that they took out behind the woodshed, I hate it.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#198

Earlier quoted context omitted.

None of the startups are profitable. What exactly are they getting beaten at? My advice is to listen less to brainrot 'influencers' that optimise for engagement through sensationalism.

Intelligence. They have "unlimited" resources and has researched AI since the very beginning - PageRank is a form of AI even. And still, Gemini is behind Claude, GPT, Grok, Muse, GLM, Kimi and is maybe on par with DeepSeek? As I said, it is embarrassing.

Their Flash model is going head to head with the SOTA models. It's completely the opposite of embarrassing.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#199
post #97

Earlier quoted context omitted.

I was thinking about coding specifically. Also, see all these math and physics breakthroughs, it's usually not Gemini but Fable and Astra.

I think 3.7 Flash is very good at coding. I have access to both that and Opus 5. Opus is only slightly better IMO and it's much more frustrating to read.

I think it's good at basic coding and is cheaper and faster. Trying it vs Fable 5.1 I don't see it being even close capability-wise. What I mean by basic coding is "write a script to parse this data" and answering some questions based on the code. It does that well and answers fast. However, when it comes to planning and thinking through a more complicated design problem it's just not there yet.

Re: Gemini 3.8 Live and 3.8 Live Extended Thinking

#200

Earlier quoted context omitted.

Have you used Google search at all on the past few months? Every single search brings up a live chat prompt. They're serving fast AI to billions of users at huge scale everyday And they're making money doing it. Perhaps they don't have the best coding model right now (although 3.8 flash is arguably SOTA at some benchmarks), but is that the be-all and end-all of AI? Only coding matters?

It’s so annoying that everyone just points to the Artificial Analysis index (or even worse, Epoch AI, where part of the score is how good the AI is at chess) as a proxy for “how good” the model is.

Its 200-300 TPS. GPT is at 50-60. And its like 5% worse? Yeah that is a good tradeoff.
Post reply on HN