Live data from Hacker News

Gemini 3.1 Pro

blog.google

151–160 of 951 posts

Re: Gemini 3.1 Pro

#151
post #82

Google seems to really pull ahead in this AI race. For me personally they offer the best deal and although the software is not quiet there compared to openai or anthropic (in regards to 1. web GUI, 2. agent-cli). I hope they can fix that in the future and I think once Gemini 4 or whatever launches we will see a huge leap again

I hope they fail.

I honestly do not wish Google to have the best model out there and be forced to use their incomprehensible subscription / billing / project management whatever shit ever again.

I don’t know what their stuff cost. I don’t know why would I use vertex or ai studio. What is included in my subscription what is billed per use.

I pray that whatever they build fails and burns.

Re: Gemini 3.1 Pro

#152

To use in OpenCode, you can update the models it has: opencode models --refresh Then /models and choose Gemini 3.1 Pro You can use the model through OpenCode Zen right away and avoid that Google UI craziness. --- It is quite pricey! Good speed and nailed all my tasks so far. For example: @app-api/app/controllers/api/availability_controller.rb @.claude/skills/healthie/SKILL.md Find Alex's id, and add him to the block…

I don't see it even after refresh. Are you using the opencode-gemini-auth plugin as well?

Re: Gemini 3.1 Pro

#153
post #92

> Last week, we released a major update to Gemini 3 Deep Think to solve modern challenges across science, research and engineering. Today, we’re releasing the upgraded core intelligence that makes those breakthroughs possible: Gemini 3.1 Pro. So this is same but not same as Gemini 3 Deep Think? Keeping track of these different releases is getting pretty ridiculous.

3.1 == model

deep think == turning up thinking knob (I think)

deep research == agent w/ search

Re: Gemini 3.1 Pro

#154

Seems like they actually fixed some of the problems with the model. Hallucinations rate seems to be much better. Seems like they also tuned the reasoning maybe that were they got most of the improvements from.

The hallucination rate with the Gemini family has always been my problem with them. Over the last year they’ve made a lot of progress catching the Gemini models up to/near the frontier in general capability and intelligence, but they still felt very late 2024 in terms of hallucination rate.

Which made the Gemini models untrustworthy for anything remotely serious, at least in my eyes. If they’ve fixed this or at least significantly improved, that would be a big deal.

Re: Gemini 3.1 Pro

#155
post #43
post #11

Earlier quoted context omitted.

The touted SVG improvements make me excited for animated pelicans.

I just gave it a shot and this is what I got: https://codepen.io/takoid/pen/wBWLOKj The model thought for over 5 minutes to produce this. It's not quite photorealistic (some parts are definitely "off"), but this is definitely a significant leap in complexity.

Good to see it wearing a helmet. Their safety team must be on their game.

Re: Gemini 3.1 Pro

#156

This model says it accepts video inputs. I asked it to transcribe a 5 second video of a digital water curtain which spelled “Boo Happy Halloween”, and it came back with “Happy” which wasn’t the first frame, but also is incomplete. This kind of test is good because it requires stitching together info from the whole video.

It reads videos at 1fps by default. You have to set the video resolution to high in ai studio

Re: Gemini 3.1 Pro

#157

Every time I've used Gemini models for anything besides code or agentic work they lean so far into the RLHF induced bold lettering and bullet point list barf that everything they output reads as if the model was talking _at_ me and not _with_ me. In my Openclaw experiment(s) and in the Gemini web UI, I've specifically added instructions to avoid this type of behavior, but it only seemed to obey those rules when I rem…

I have no issues adjusting gemini tone & style with system prompt content

Re: Gemini 3.1 Pro

#158
post #52

Pretty great pelican: https://simonwillison.net/2026/Feb/19/gemini-31-pro/ - took over 5 minutes though, but I think that's because they're having performance teething problems on launch day.

It's an excellent demonstration of the main issue I have with the Gemini family of models, they always go "above and beyond" to do a lot of stuff, even if I explicitly prompt against it. In this case, most of the SVG ends up consisting not just of a bike and a pelican, but clouds, a sun, a hat on the pelican and so much more. Exactly the same thing happens when you code, it's almost impossible to get Gemini to not do…

> it's almost impossible to get Gemini to not do "helpful" drive-by-refactors

Just asking "Explain what this service does?" turns into

[No response for three minutes...]

+729 -522

Re: Gemini 3.1 Pro

#159

I really want to use google’s models but they have the classic Google product problem that we all like to complain about. I am legit scared to login and use Gemini CLI because the last time I thought I was using my “free” account allowance via Google workspace. Ended up spending $10 before realizing it was API billing and the UI was so hard to figure out I gave up. I’m sure I can spend 20-40 more mins to sort this ou…

You could always use it through Copilot. The credits based billing is pretty simple without surprise charges.

Re: Gemini 3.1 Pro

#160

Earlier quoted context omitted.

Models are soon going to start benchmaxxing generating SVGs of pelicans on bikes

Simons been doing this exact test for nearly 18 months now, if vendors want to benchmaxx it then they've had more than enough time to do so already.

Exactly. As far as I'm concerned, the benchmark is useless. It's way too easy and rewarding to train on it.
Post reply on HN