Live data from Hacker News

Gemini 3.1 Pro

blog.google

691–700 of 951 posts

Re: Gemini 3.1 Pro

#691
post #675

Earlier quoted context omitted.

You can’t put Gemini and Meta in the same sentence. Llama 4 was DOA, and Meta has given up on frontier models. Internally they’re using Claude.

After spending all that money and firing a bunch of people? Is the new group doing anything at this point?

They are busy demonstrating that Mark Zuckerberg has no sense at all.

Re: Gemini 3.1 Pro

#692
post #559

Earlier quoted context omitted.

While price is definitely important, results are extremely important. Gemini often falls into the 'didn't do' it part of the spectrum, this days Opus almost always does 'good enough'. Gemini definitely has its merits but for me it just doesn't do what other models can. I vibe-coded an app which recommends me restaurants. The app uses gemini API to make restaurants given bunch of data and prompt. App itself is vibe-co…

The binary you draw on models that havent been out a quarter is borderline insane. Opus is absurdly good in Claude code but theres a lot of use cases Gemini is great at. I think Google is further behind with the harness than the model

I was careful not to draw binary. I was saying that Opus in Claude Code is good enough for me to make projects. Using Gemini after it seems like a significant downgrade, which actually doesn't get the job done helping me code. This is my experience, it can change if Gemini will get better.

However, for internal use I opt to Gemini, because of API cost. It is great in sorting reviews and menues out.

Re: Gemini 3.1 Pro

#693

Earlier quoted context omitted.

Spark is fun and cool, but it isn't some revolution. It's a different workflow, but not suitable for everything that you're use GPT5.2 for with thinking set to high, for example, it's way more dumb and makes more mistakes, while 5.2 will carefully thread through a large codebase and spend 40 minutes just to validate the change actually didn't break anything, as long as you provide prompts for it. Spark on the other h…

Spark is the 'same model and harness' but on Cerebras. Your intuition may be deceiving you, maybe assuming it's a speed/quality trade-off, it's not. It's just faster hardware. No IQ tradeoff. If you toy around with Cerebras directly, you get a feel for it. Edit: see note below, I'm wrong. Not same model.

> Today, we’re releasing a research preview of GPT‑5.3‑Codex‑Spark, a smaller version of GPT‑5.3‑Codex, and our first model designed for real-time coding.

from https://openai.com/index/introducing-gpt-5-3-codex-spark/, emphasis mine

Re: Gemini 3.1 Pro

#694
post #522
post #506

Earlier quoted context omitted.

They've fallen way behind.

GPT 5.2 loses at everything but they included that

Who are they supposed to compare it to? I'm not sure what makes you think that Grok is even remotely comparable to the frontier models right now.

Re: Gemini 3.1 Pro

#695

Earlier quoted context omitted.

> I still have no idea how to pay for Gemini CLI, in codex/claude its very simple $20/month for entry and $200/month for ton of weekly usage. This! I would like to sign up for a paid plan for Gemini CLI. But I have not been able to figure out how. I already have Codex and Claude plans. Those were super easy to sign up for.

What’s your difficulty? Google has published easy to follow 27-step instructions for how to sign up for the half a dozen services you need to chain together to enable this common usecase!

On the 3.0 rollout I signed up for billing and it just silently failed. Solution was to remake billing account and then wait a day

Re: Gemini 3.1 Pro

#696

People underrate Google's cost effectiveness so much. Half price of Opus. HALF. Think about ANY other product and what you'd expect from the competition thats half the price. Yet people here act like Gemini is dead weight ____ Update: 3.1 was 40% of the cost to run AA index vs Opus Thinking AND SONNET, beat Opus, and still 30% faster for output speed. https://artificialanalysis.ai/?speed=intelligence-vs-speed&m...

If something is shit, it doesn't matter it costs half price of something okay.

"There is hardly anything in the world that some man cannot make a little worse and sell a little cheaper, and the people who consider price only are this man's lawful prey."

Re: Gemini 3.1 Pro

#697

Earlier quoted context omitted.

Yes, this is very true and it speaks strongly to this wayward notion of 'models' - it depends so much on the tuning, the harness, the tools. I think it speaks to the broader notion of AGI as well. Claude is definitively trained on the process of coding not just the code, that much is clear. Codex has the same limitation but not quite as bad. This may be a result of Anthropic using 'user cues' with respect to what are…

Google are stuck because they have to compete with OpenAI. If they don’t, they face an existential threat to their advertising business. But then they leave the door open for Anthropic on coding, enterprise and agentic workflows. Sensibly, that’s what they seem to be doing. That said Gemini is noticeably worse than ChatGPT (it’s quite erratic) and Anthropic’s work on coding / reasoning seems to be filtering back to i…

Google might be a mess now, but they have time. OpenAI and Anthropic are on barrowed time, Google has a built in money printer. They just need to outlast the others.

Re: Gemini 3.1 Pro

#698
post #141
post #52

Pretty great pelican: https://simonwillison.net/2026/Feb/19/gemini-31-pro/ - took over 5 minutes though, but I think that's because they're having performance teething problems on launch day.

Wonder when will we get something other than a side view

Another Jeff Dean post about this model shows it writing programs that generate CAD objects. I suspect if you ask it to, it will create a CAD pelican on a CAD bicycle and even make joints so you can turn the pedals.

Re: Gemini 3.1 Pro

#699
post #637
post #306

These models are so powerful. It's totally possible to build entire software products in the fraction of the time it took before. But, reading the comments here, the behaviors from one version to another point version (not major version mind you) seem very divergent. It feels like we are now able to manage incredibly smart engineers for a month at the price of a good sushi dinner. But it also feels like you have to b…

I keep giving the top Anthropic, Google and OpenAI models problems. They come up with passable solutions and are good for getting juices flowing and giving you a start on a codebase, but they are far from building "entire software products" unless you really don't care about quality and attention to detail.

That is my experience too. I don't know what others are building but the more novel the task is the worse these models perform.

Re: Gemini 3.1 Pro

#700

Gemini 3 is still in preview (limited rate limits) and 2.5 is deprecated (still live but won't be for long).[0] Are Google planning to put any of their models into production any time soon? Also somewhat funny that some models are deprecated without a suggested alternative(gemini-2.5-flash-lite). Do they suggest people switch to Claude? [0] https://ai.google.dev/gemini-api/docs/deprecations

They probably have some inflexible internal policy where preview needs to be in use for X months before GA. Couple that with the rate of AI progress and voila.
Post reply on HN