Live data from Hacker News

Gemini 3 Pro Model Card [pdf]

storage.googleapis.com

271–280 of 359 posts

Re: Gemini 3 Pro Model Card [pdf]

#271

These model cards tell me nothing. I want to know the exact data a model was trained on. Otherwise, how can I safely use it for generating texts that I show to children? Etc.etc.

Shouldn't you be carefully reading texts before you show it to children?

Re: Gemini 3 Pro Model Card [pdf]

#272
post #74
post #70

Earlier quoted context omitted.

Looks like it will be on par with the contenders when it comes to coding. I guess improvements will be incremental from here on out.

If it’s on par in code quality, it would be a way better model for coding because of its huge context window.

Sonnet can also work on 1M context. Its extreme speed is the only thing Gemini has on others.

Re: Gemini 3 Pro Model Card [pdf]

#273

Earlier quoted context omitted.

Creator of pixeldrain here. Italy has been doing this for a very long time. They never notified me of any such material being present on my site. I have a lot of measures in place to prevent the spread of CSAM. I have sent dozens of mails to Polizia Postale and even tried calling them a few times, but they never respond. My mails go unanswered and they just hang up the phone.

Have you tried Europol?

Not yet. I also thought about reaching out to the embassy, but have not had the time for it yet.

Re: Gemini 3 Pro Model Card [pdf]

#274

Earlier quoted context omitted.

Not to mention no macOS app. This is probably unimportant to many in the hn audience, but more broadly it matters for your average knowledge worker.

And a REALLY good macOS app. Like, kind of unreasonably good. You’d expect some perfunctory Electronic app that just barely wraps the website. But no, you get something that feels incredibly polished…more so than a lot of recent apps from Apple…and has powerful integrations into other apps, including text editors and terminals.

Which app are you referring to?

Re: Gemini 3 Pro Model Card [pdf]

#275

Earlier quoted context omitted.

Have you tried Europol?

Not yet. I also thought about reaching out to the embassy, but have not had the time for it yet.

As far as I know, Europol can route your report to appropriate local authority.

Re: Gemini 3 Pro Model Card [pdf]

#276
post #98

Earlier quoted context omitted.

These numbers are impressive, at least to say. It looks like Google has produced a beast that will raise the bar even higher. What's even more impressive is how Google came into this game late and went from producing a few flops to being the leader at this (actually, they already achieved the title with 2.5 Pro). What makes me even more curious is the following > Model dependencies: This model is not a modification o…

At least at the moment, coming in late seems to matter little. Anyone with money can trivially catch up to a state of the art model from six months ago. And as others have said, late is really a function of spigot, guardrails, branding, and ux, as much as it is being a laggard under the hood.

One possibility here is that Google is dribbling out cutting edge releases to slowly bleed out the pure play competition.

Re: Gemini 3 Pro Model Card [pdf]

#277

Earlier quoted context omitted.

Not yet. I also thought about reaching out to the embassy, but have not had the time for it yet.

As far as I know, Europol can route your report to appropriate local authority.

Thanks, I'll give them a call tomorrow. The website only lists a dutch phone number, which is convenient, I'm dutch as well.

Re: Gemini 3 Pro Model Card [pdf]

#278
post #272
post #74

Earlier quoted context omitted.

If it’s on par in code quality, it would be a way better model for coding because of its huge context window.

Sonnet can also work on 1M context. Its extreme speed is the only thing Gemini has on others.

Can it now in Claude Code and Claude Desktop? When I was using it a couple of months ago it seemed only the API had 1M

Re: Gemini 3 Pro Model Card [pdf]

#279
post #75

It is interesting that the Gemini 3 beats every other model on these benchmarks, mostly by a wide margin, but not on SWE Bench. Sonnet is still king here and all three look to be basically on the same level. Kind of wild to see them hit such a wall when it comes to agentic coding

Their scores on SWE bench are very close because the benchmark is nearly saturated. Gemini 3 beats Sonnet 4.5 on TerminalBench 2.0 by a nice margin (54% vs. 43%), which is also agentic coding (CLI instead of python).

Re: Gemini 3 Pro Model Card [pdf]

#280
post #65

Earlier quoted context omitted.

Anthropic has a fairly significant lead when it comes to enterprise usage and for coding. This seems like a workable business model to me.

I feel this is a tenuous position though. I find it incredibly easy to switch to Gemini CLI when I want a second opinion, or when Claude is down.

The enterprise sales cycle is often quite long, though, and often includes a lot of hurdles around compliance, legal, etc. It would take a fairly sustained loss of edge before a lot of enterprises would switch once they're hooked into a given platform. It's interesting to me that Sonnet 4.5 still edges Gemini 3 on SWE bench. This seems to bode well for the trajectory that Anthropic is on.
Post reply on HN