There needs to be a sycophancy benchmark in these comparisons. More baseless praise and false agreement = lower score.
Gemini 3 Pro Model Card [pdf]
61–70 of 359 posts
Re: Gemini 3 Pro Model Card [pdf]
#62[1] https://blog.google/technology/ai/introducing-pathways-next-...
Re: Gemini 3 Pro Model Card [pdf]
#63Re: Gemini 3 Pro Model Card [pdf]
#64 This model is not a modification or a fine-tune of a prior model
Is that common to mention that? Feels like they built something from scratchRe: Gemini 3 Pro Model Card [pdf]
#65Earlier quoted context omitted.
While I don’t disagree that Google is the company you can’t bet against when it comes to AI, saying other companies are done is a stretch. If they have a significant moat then they should be at the top all the time by then which is not the case though.
Agreed, too early to write off others entirely. It'll be interesting to see who comes out the other side of the bubble with a working business.
Re: Gemini 3 Pro Model Card [pdf]
#66It says it's been trained from scratch. I wonder if it will have the same undescribable magic that makes me spend an hour every day with 2.5. I really love the results I can get with 2.5 pro. Google eventually limiting aistudio will be a sad day. Also I really hoped for a 2M+ context. I'm living on the context edge even with 1M.
Re: Gemini 3 Pro Model Card [pdf]
#67Earlier quoted context omitted.
Why? These models just leapfrog each other as time advances. One month Gemini is on top, then ChatGPT, then Anthropic. Not sure why everyone gets FOMO whenever a new version gets released.
Considering GPT 5 was only recently released, it's very unlikely GPT will achieve these scores in just a couple of months. If they had something this good in the oven, they'd probably left the GPT 5 name to it. Or maybe Google just benchmaxxed and this doesn't translate at all in real world performance.
Re: Gemini 3 Pro Model Card [pdf]
#68If these numbers are true then OpenAI is probably done, Anthropic too. Still, it's hard to see an effective monetization method for this tech and it clearly is eating Google's main pie which is search.
1) New SOTA models come out all the time and that hasn't killed the other major AI companies. This will be no different. 2) Google's search revenue last quarter was $56 billion, a 14% increase over Q3 2024.
2) I'm not suggesting this will happen overnight but especially younger people gravitate towards LLM for information search + actively use some sort of ad blocking. In the long run it doesn't look great for Google.
Re: Gemini 3 Pro Model Card [pdf]
#69Title of the document is "[Gemini 3 Pro] External Model Card - November 18, 2025 - v2", in case you needed further confirmation that the model will be released today. Also interesting to know that Google Antigravity (antigravity.google / https://github.com/Google-Antigravity ?) leaked. I remember seeing this subdomain recently. Probably Gemini 3 related as well. Org was created on 2025-11-04T19:28:13Z ( https://api.g…
Re: Gemini 3 Pro Model Card [pdf]
#70Benchmarks from page 4 of the model card: | Benchmark | 3 Pro | 2.5 Pro | Sonnet 4.5 | GPT-5.1 | |-----------------------|-----------|---------|------------|-----------| | Humanity's Last Exam | 37.5% | 21.6% | 13.7% | 26.5% | | ARC-AGI-2 | 31.1% | 4.9% | 13.6% | 17.6% | | GPQA Diamond | 91.9% | 86.4% | 83.4% | 88.1% | | AIME 2025 | | | | | | (no tools) | 95.0% | 88.0% | 87.0% | 94.0% | | (code execution) | 100% | -…