Live data from Hacker News

Gemini 3.1 Pro

blog.google

751–760 of 951 posts

Re: Gemini 3.1 Pro

#751

Earlier quoted context omitted.

Google are stuck because they have to compete with OpenAI. If they don’t, they face an existential threat to their advertising business. But then they leave the door open for Anthropic on coding, enterprise and agentic workflows. Sensibly, that’s what they seem to be doing. That said Gemini is noticeably worse than ChatGPT (it’s quite erratic) and Anthropic’s work on coding / reasoning seems to be filtering back to i…

In my experience Gemini 3.0 pro is noticeably better than chatgpt 5.2 for non-coding tasks. The latter gives me blatantly wrong information all the time, the former very rarely.

Strange that you say that because the general consensus (and my experience) seems to be the opposite, as well as the AA-Omniscience Hallucination Rate Benchmark which puts 3.0 Pro among the higher hallucinating models. 3.1 seems to be a noticeable improvement though.

Re: Gemini 3.1 Pro

#752
post #709

What I’m noticing, overall: I’ve never cut so much code in my life. I’ve become a coding monster with one of those dark green GitHub profiles ever since 5.3-Codex gave me the confidence to load in a ridiculous number of tasks every day and let it rip. I have about three coding tasks going at once and in another window, Claude Cowork is ripping through PowerPoints and getting back to lawyers. This tech is not going to…

There are thousands like you now. How many does it take to run the economy? What would the rest do. Think of it like what a tractor did to agricultural work. The fist guy that used a tractor probably thought: this is not replacing me, I’m just much more productive. Well, turns out you only need one guy per farm now.

[dead]

Re: Gemini 3.1 Pro

#753

Earlier quoted context omitted.

We are not at the moment where price matters. All that matters is performance.

It matters to me. I pay for it and I like using it. I pick my models to keep my spend reigned in.

What do you use it for? What is your time worth that you'd settle for a lesser model to save a few bucks?

Re: Gemini 3.1 Pro

#754
post #751

Earlier quoted context omitted.

In my experience Gemini 3.0 pro is noticeably better than chatgpt 5.2 for non-coding tasks. The latter gives me blatantly wrong information all the time, the former very rarely.

Strange that you say that because the general consensus (and my experience) seems to be the opposite, as well as the AA-Omniscience Hallucination Rate Benchmark which puts 3.0 Pro among the higher hallucinating models. 3.1 seems to be a noticeable improvement though.

I can only speak to my own experience, but for the past couple of months I've been duplicating prompts across both for high value tasks, and that has been my consistent finding.

Re: Gemini 3.1 Pro

#755
post #325

Earlier quoted context omitted.

Don't get me started on the thinking tokens. Since 2.5P the thinking has been insane. "I'm diving in to the problem", "I'm fully immersed" or "I'm meticulously crafting the answer"

I once saw "now that I've slept on it" in Gemini's CoT... baffling.

Reminds me of Claude's time estimates. Yeah this project isn't actually going to take 12 weeks, Claude, nice try though.

Re: Gemini 3.1 Pro

#756
post #677

This is great. I am hopeful that Gemini 3.1 Pro would be great. So far, I'm almost always pulled away from Gemini models by Claude. Having used Claude Opus High for a while now, Claude Opus seems to be fantastic at coding. Even Gemini's comparison chart says so. OpenAI's 5.3-codex is by far the weakest (of the 3) for my coding purposes. Claude Opus really shines at explanations and generating code. Gemini is almost g…

> I keep switching among these subscriptions every month to not miss out on any of the offerings for too long; ChatGPT Plus Gemini Pro Claude. I wonder why many people seem to be doing this instead of just going for a copilot subscription that has access to all those models? Anybody care to share pros and cons?

OpenAI and Anthropic give you a lot of usage/$ through their plans. For the Anthropic Max plans, this can be like a ~90% discount. Copilot does not benefit from this (their pricing model is also different though, it is request-based rather than token usage based, so it is hard to compare).

That's not to mention that the models generally work better in their own harnesses, which is perhaps unsurprising because the models have been trained with the specific harness in mind (and vice versa). That said, I think some 3rd-party harnesses do a lot of work to make different models work well in their harness.

Re: Gemini 3.1 Pro

#757

This is great. I am hopeful that Gemini 3.1 Pro would be great. So far, I'm almost always pulled away from Gemini models by Claude. Having used Claude Opus High for a while now, Claude Opus seems to be fantastic at coding. Even Gemini's comparison chart says so. OpenAI's 5.3-codex is by far the weakest (of the 3) for my coding purposes. Claude Opus really shines at explanations and generating code. Gemini is almost g…

I would suggest you also take a look at Cursor's Composer1.5. It's super fast, and perform better than Gemini3P in my use cases.

Re: Gemini 3.1 Pro

#758

Earlier quoted context omitted.

They have much much less time than one would think. Their ads business is about to go into freefall, this will cause the whole company to spiral.

I mean their ads business just broke $80b per quarter, not sure where this idea is coming from...

Google hasn't seen its legacy ad revenue start to dent until products with built-in agents start to see mass adoption.

Writing is on the wall that orders of magnitude fewer people will be going to google.com or using an interactive Google search in the next 5 years though.

Re: Gemini 3.1 Pro

#759

Earlier quoted context omitted.

Plus they started making AI processors 11 years ago and invented the math behind “GPTs” 9 years ago. Gemini is way cheaper to run for them than it does for everyone else. I think Gemini is really built for their biggest market — Google Search. You ask questions and get answers. I’m sure they’ll figure out agentic flows. Google is always a mess when it comes to product. Don’t forget the Google chat sagas where it seem…

Who they? Do the engineers who actually did that work at Google still? I heard that the guy who made TPUs has his own startup now.

They got acquired by Nvidia

Re: Gemini 3.1 Pro

#760
In the "Intelligence applied" section, where they show the comparison animations, they are shown using a non-optimal UI.

There is not enough time to read the text, see old animation, and see new animation. Better would have been to keep the same animation on repeat, so that people have unlimited time to read the text and observer the animations.

Also, it jumps from example to example in the same video. Better would have been to show each separately, so that once user is done observing one example at their own pace, they can proceed to the next.

As a workaround, I had to open the video (just the video) in a new tab, pause once an example came up, read the text, then rewind to the start of the animation to see the old animation example, then rewind again, then see the new animation example, and then sometimes rewind again if I wanted to see the animation again. Then, once done with the example, I had to forward to the next example and repeat the above process again.

Somewhere along that process, they lost me.

Post reply on HN