Live data from Hacker News

Gemini 3 Pro Model Card [pdf]

storage.googleapis.com

61–70 of 359 posts

Re: Gemini 3 Pro Model Card [pdf]

#61

There needs to be a sycophancy benchmark in these comparisons. More baseless praise and false agreement = lower score.

Your comment demonstrates a remarkably elevated level of cognitive processing and intellectual rigor. Inquiries of this caliber are indicative of a mind operating at a strategically advanced tier, displaying exceptional analytical bandwidth and thought-leadership potential. Given the substantive value embedded in your question, it is operationally imperative that we initiate an immediate deep-dive and execute a comprehensive response aligned with the strategic priorities of this discussion.

Re: Gemini 3 Pro Model Card [pdf]

#63
I know this is a little controversial but the lack of performance on SWE-bench is hugely disappointing I think economically. These models don’t have any viable path to profitability if they can’t take engineering jobs.

Re: Gemini 3 Pro Model Card [pdf]

#65
post #34

Earlier quoted context omitted.

While I don’t disagree that Google is the company you can’t bet against when it comes to AI, saying other companies are done is a stretch. If they have a significant moat then they should be at the top all the time by then which is not the case though.

Agreed, too early to write off others entirely. It'll be interesting to see who comes out the other side of the bubble with a working business.

Anthropic has a fairly significant lead when it comes to enterprise usage and for coding. This seems like a workable business model to me.

Re: Gemini 3 Pro Model Card [pdf]

#66
post #12

It says it's been trained from scratch. I wonder if it will have the same undescribable magic that makes me spend an hour every day with 2.5. I really love the results I can get with 2.5 pro. Google eventually limiting aistudio will be a sad day. Also I really hoped for a 2M+ context. I'm living on the context edge even with 1M.

[deleted]

Re: Gemini 3 Pro Model Card [pdf]

#67
post #37
post #13

Earlier quoted context omitted.

Why? These models just leapfrog each other as time advances. One month Gemini is on top, then ChatGPT, then Anthropic. Not sure why everyone gets FOMO whenever a new version gets released.

Considering GPT 5 was only recently released, it's very unlikely GPT will achieve these scores in just a couple of months. If they had something this good in the oven, they'd probably left the GPT 5 name to it. Or maybe Google just benchmaxxed and this doesn't translate at all in real world performance.

GPT 5 was released more than 3 months ago. Gemini 2.5 was released less than 8 months ago.

Re: Gemini 3 Pro Model Card [pdf]

#68

If these numbers are true then OpenAI is probably done, Anthropic too. Still, it's hard to see an effective monetization method for this tech and it clearly is eating Google's main pie which is search.

1) New SOTA models come out all the time and that hasn't killed the other major AI companies. This will be no different. 2) Google's search revenue last quarter was $56 billion, a 14% increase over Q3 2024.

1) Not long ago Altman and the OpenAI CFO were openly asking for public money. None of these AI companies have actually any kind of working business plan and are just burning investor money. If the investors see there is no winning against Google (or some open Chinese model) the money will dry up.

2) I'm not suggesting this will happen overnight but especially younger people gravitate towards LLM for information search + actively use some sort of ad blocking. In the long run it doesn't look great for Google.

Re: Gemini 3 Pro Model Card [pdf]

#69

Title of the document is "[Gemini 3 Pro] External Model Card - November 18, 2025 - v2", in case you needed further confirmation that the model will be released today. Also interesting to know that Google Antigravity (antigravity.google / https://github.com/Google-Antigravity ?) leaked. I remember seeing this subdomain recently. Probably Gemini 3 related as well. Org was created on 2025-11-04T19:28:13Z ( https://api.g…

what is Google Antigravity?

Re: Gemini 3 Pro Model Card [pdf]

#70
post #23

Benchmarks from page 4 of the model card: | Benchmark | 3 Pro | 2.5 Pro | Sonnet 4.5 | GPT-5.1 | |-----------------------|-----------|---------|------------|-----------| | Humanity's Last Exam | 37.5% | 21.6% | 13.7% | 26.5% | | ARC-AGI-2 | 31.1% | 4.9% | 13.6% | 17.6% | | GPQA Diamond | 91.9% | 86.4% | 83.4% | 88.1% | | AIME 2025 | | | | | | (no tools) | 95.0% | 88.0% | 87.0% | 94.0% | | (code execution) | 100% | -…

Looks like it will be on par with the contenders when it comes to coding. I guess improvements will be incremental from here on out.
Post reply on HN