I've been using the preview flash model exclusively since it came out, the speed and quality of response is all I need at the moment. Although still using Claude Code w/ Opus 4.5 for dev work. Google keeps their models very "fresh" and I tend to get more correct answers when asking about Azure or O365 issues, ironically copilot will talk about now deleted or deprecated features more often.
Gemini 3 Flash: Frontier intelligence built for speed
81–90 of 609 posts
Re: Gemini 3 Flash: Frontier intelligence built for speed
#82Re: Gemini 3 Flash: Frontier intelligence built for speed
#83Wild how this beats 2.5 Pro in every single benchmark. Don't think this was true for Haiku 4.5 vs Sonnet 3.5.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#84Re: Gemini 3 Flash: Frontier intelligence built for speed
#85You can get your HN profile analyzed and roasted by it. It's pretty funny :) https://hn-wrapped.kadoa.com
Re: Gemini 3 Flash: Frontier intelligence built for speed
#86It has a SimpleQA score of 69%, a benchmark that tests knowledge on extremely niche facts, that's actually ridiculously high (Gemini 2.5 *Pro* had 55%) and reflects either training on the test set or some sort of cracked way to pack a ton of parametric knowledge into a Flash Model. I'm speculating but Google might have figured out some training magic trick to balance out the information storage in model capacity. Tha…
More experts with a lower pertentage of active ones -> more sparsity.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#87This is the first flash/mini model that doesn't make a complete ass of itself when I prompt for the following: "Tell me as much as possible about Skatval in Norway. Not general information. Only what is uniquely true for Skatval." Skatval is a small local area I live in, so I know when it's bullshitting. Usually, I get a long-winded answer that is PURE Barnum-statement, like "Skatval is a rural area known for its bea…
Re: Gemini 3 Flash: Frontier intelligence built for speed
#88Does anyone else understand what the difference is between Gemini 3 'Thinking' and 'Pro'? Thinking "Solves complex problems" and Pro "Thinks longer for advanced math & code". I assume that these are just different reasoning levels for Gemini 3, but I can't even find mention of there being 2 versions anywhere, and the API doesn't even mention the Thinking-Pro dichotomy.
- "Thinking" is Gemini 3 Flash with higher "thinking_level"
- Prop is Gemini 3 Pro. It doesn't mention "thinking_level" but I assume it is set to high-ish.Re: Gemini 3 Flash: Frontier intelligence built for speed
#89I’m wondering why Claude Opus 4.5 is missing from the benchmarks table.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#90Even before this release the tools (for me: Claude Code and Gemini for other stuff) reached a "good enough" plateau that means any other company is going to have a hard time making me (I think soon most users) want to switch. Unless a new release from a different company has a real paradigm shift, they're simply sufficient. This was not true in 2023/2024 IMO. With this release the "good enough" and "cheap enough" int…
For me, the last wave of models finally started delivering on their agentic coding promises.