Live data from Hacker News

Gemini 3 Flash: Frontier intelligence built for speed

blog.google

11–20 of 609 posts

Re: Gemini 3 Flash: Frontier intelligence built for speed

#11
post #3

They went too far, now the Flash model is competing with their Pro version. Better SWE-bench, better ARC-AGI 2 than 3.0 Pro. I imagine they are going to improve 3.0 Pro before it's no more in Preview. Also I don't see it written in the blog post but Flash supports more granular settings for reasoning: minimal, low, medium, high (like openai models), while pro is only low and high.

> They went too far, now the Flash model is competing with their Pro version

Wasn't this the case with the 2.5 Flash models too? I remember being very confused at that time.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#12
post #7

These flash models keep getting more expensive with every release. Is there an OSS model that's better than 2.0 flash with similar pricing, speed and a 1m context window? Edit: this is not the typical flash model, it's actually an insane value if the benchmarks match real world usage. > Gemini 3 Flash achieves a score of 78%, outperforming not only the 2.5 series, but also Gemini 3 Pro. It strikes an ideal balance fo…

cost of e2e task resolution should be cheaper, even if single inference cost is higher, you need fewer loops to solve a problem now

Re: Gemini 3 Flash: Frontier intelligence built for speed

#16
post #3

They went too far, now the Flash model is competing with their Pro version. Better SWE-bench, better ARC-AGI 2 than 3.0 Pro. I imagine they are going to improve 3.0 Pro before it's no more in Preview. Also I don't see it written in the blog post but Flash supports more granular settings for reasoning: minimal, low, medium, high (like openai models), while pro is only low and high.

I'm not sure how I'm going to live with this!

Re: Gemini 3 Flash: Frontier intelligence built for speed

#18

Even before this release the tools (for me: Claude Code and Gemini for other stuff) reached a "good enough" plateau that means any other company is going to have a hard time making me (I think soon most users) want to switch. Unless a new release from a different company has a real paradigm shift, they're simply sufficient. This was not true in 2023/2024 IMO. With this release the "good enough" and "cheap enough" int…

For me, the last wave of models finally started delivering on their agentic coding promises.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#19

Two quick questions to Gemini/AI Studio users: 1, has anyone actually found 3 Pro better than 2.5 (on non code tasks)? I struggle to find a difference beyond the quicker reasoning time and fewer tokens. 2, has anyone found any non-thinking models better than 2.5 or 3 Pro? So far I find the thinking ones significantly ahead of non thinking models (of any company for that matter.)

Gemini 3 is a step change up against 2.5 for electrical engineering R&D.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#20
From the article, speed & cost match 2.5 Flash. I'm working on a project where there's a huge gap between 2.5 Flash and 2.5 Flash Lite as far as performance and cost goes.

-> 2.5 Flash Lite is super fast & cheap (~1-1.5s inference), but poor quality responses.

-> 2.5 Flash gives high quality responses, but fairly expensive & slow (5-7s inference)

I really just need an in-between for Flash and Flash Lite for cost and performance. Right now, users have to wait up to 7s for a quality response.

Post reply on HN