Gemini 3.8 Live and 3.8 Live Extended Thinking
191–200 of 334 posts
Re: Gemini 3.8 Live and 3.8 Live Extended Thinking
#192I wonder when/if we’ll see Gemini beating Fable and Astra. Last year I would have confidently bet Google will overtake the others just because they have the data, the hardware (TPUs) and a fat advertising money pipe and yet they are still behind. Anyone anonymous at Google want to hint when Gemini 4 will be out?
I’m wondering if they even see a coding agent as a valuable prize. It’s a competitive market in a race to the bottom economically, hard to establish consistent differentiation and virtually zero switching cost for customers. I think they’ve made a shrewd move in focusing on search integration and everyday users (Gemini app) vs software power users. They have their corner and nobody is really competing with them, plus…
Maybe everybody is wrong about how valuable these companies will be, but atm it looks very dumb to not be competitive in coding.
Re: Gemini 3.8 Live and 3.8 Live Extended Thinking
#193Earlier quoted context omitted.
Intelligence. They have "unlimited" resources and has researched AI since the very beginning - PageRank is a form of AI even. And still, Gemini is behind Claude, GPT, Grok, Muse, GLM, Kimi and is maybe on par with DeepSeek? As I said, it is embarrassing.
If RSI is achievable, it will leapfrog everything produced so far and so it will make sense to focus on RSI instead of incremental improvements for your top model. Startups need investment and need to show progress. Google does not at the moment need to take lead in the current race.
Re: Gemini 3.8 Live and 3.8 Live Extended Thinking
#194My first language is Afrikaans, which is a somewhat niche language and hard to find teachers/conversation buddies outside South Africa. (I live in USA now) I've been using Gemini to live chat in Afrikaans and do impromptu Afrikaans grammar lessons during my solo drives around town. It is phenomenal at speaking the language - like, it really shocks my family members when they hear it. This is probably the most joy I g…
I love Gemini in Google Maps for long drives. I start getting bored of music and podcasts and start grilling it with random questions I've always wondered about.
Re: Gemini 3.8 Live and 3.8 Live Extended Thinking
#195Re: Gemini 3.8 Live and 3.8 Live Extended Thinking
#196Earlier quoted context omitted.
I’m curious how you guys keep track of each model’s coding capabilities. The landscape keeps changing. I don’t suppose you benchmark all frontier models every other month, right?
I use them. Daily. Gemini hasn’t been a contender by comparison for a long time.
Re: Gemini 3.8 Live and 3.8 Live Extended Thinking
#197Re: Gemini 3.8 Live and 3.8 Live Extended Thinking
#198Earlier quoted context omitted.
None of the startups are profitable. What exactly are they getting beaten at? My advice is to listen less to brainrot 'influencers' that optimise for engagement through sensationalism.
Intelligence. They have "unlimited" resources and has researched AI since the very beginning - PageRank is a form of AI even. And still, Gemini is behind Claude, GPT, Grok, Muse, GLM, Kimi and is maybe on par with DeepSeek? As I said, it is embarrassing.
Re: Gemini 3.8 Live and 3.8 Live Extended Thinking
#199Earlier quoted context omitted.
I was thinking about coding specifically. Also, see all these math and physics breakthroughs, it's usually not Gemini but Fable and Astra.
I think 3.7 Flash is very good at coding. I have access to both that and Opus 5. Opus is only slightly better IMO and it's much more frustrating to read.
Re: Gemini 3.8 Live and 3.8 Live Extended Thinking
#200Earlier quoted context omitted.
Have you used Google search at all on the past few months? Every single search brings up a live chat prompt. They're serving fast AI to billions of users at huge scale everyday And they're making money doing it. Perhaps they don't have the best coding model right now (although 3.8 flash is arguably SOTA at some benchmarks), but is that the be-all and end-all of AI? Only coding matters?
It’s so annoying that everyone just points to the Artificial Analysis index (or even worse, Epoch AI, where part of the score is how good the AI is at chess) as a proxy for “how good” the model is.