Live data from Hacker News

Gemini 3 Flash: Frontier intelligence built for speed

blog.google

141–150 of 609 posts

Re: Gemini 3 Flash: Frontier intelligence built for speed

#141

I think about what would be most terrifying to Anthropic and OpenAI i.e. The absolute scariest thing that Google could do. I think this is it: Release low latency, low priced models with high cognitive performance and big context window, especially in the coding space because that is direct, immediate, very high ROI for the customer. Now, imagine for a moment they had also vertically integrated the hardware to do thi…

"Now, imagine for a moment they had also vertically integrated the hardware to do this."

Then you realise you aren't imagining it.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#143
post #136
post #125

At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…

"At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed" This has been true for at least 4 months and yeah, based on how these things scale and also Google's capital + in-house hardware advantages, it's probably insurmountable.

Yeah the only thing standing in Google's way is Google. And it's the easy stuff, like sensible billing models, easy to use docs and consoles that make sense and don't require 20 hours to learn/navigate, and then just the slew of bugs in Gemini CLI that are basic usability and model API interaction things. The only differentiator that OpenAI still has is polish.

Edit: And just to add an example: openAI's Codex CLI billing is easy for me. I just sign up for the base package, and then add extra credits which I automatically use once I'm through my weekly allowance. With Gemini CLI I'm using my oauth account, and then having to rotate API keys once I've used that up.

Also, Gemini CLI loves spewing out its own chain of thought when it gets into a weird state.

Also Gemini CLI has an insane bias to action that is almost insurmountable. DO NOT START THE NEXT STAGE still has it starting the next stage.

Also Gemini CLI has been terrible at visibility on what it's actually doing at each step - although that seems a bit improved with this new model today.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#144
post #106
post #57

Ok, I was a bit addicted to Opus 4.5 and was starting to feel like there's nothing like it. Turns out Gemini 3 Flash is pretty close. The Gemini CLI is not as good but the model more than makes up for it. The weird part is Gemini 3 Pro is nowhere as good an experience. Maybe because its just so slow.

I will have to try that. Cursor bill got pretty high with Opus 4.5. Never considered opus before the 4.5 price drop but now it's hard to change... :)

$100 Claude max is the best subscription I’ve ever had.

Well worth every penny now

Re: Gemini 3 Flash: Frontier intelligence built for speed

#145
post #7

These flash models keep getting more expensive with every release. Is there an OSS model that's better than 2.0 flash with similar pricing, speed and a 1m context window? Edit: this is not the typical flash model, it's actually an insane value if the benchmarks match real world usage. > Gemini 3 Flash achieves a score of 78%, outperforming not only the 2.5 series, but also Gemini 3 Pro. It strikes an ideal balance fo…

For my apps evals Gemini flash and grok 4 fast are the only ones worth using. I'd love for an open weights model to compete in this arena but I haven't found one.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#146
post #125

At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…

GPT 5.2 is actually getting me better outputs than Opus 4.5 on very complex reviews (on high, I never use less) - but the speed makes Opus the default for 95% of use cases.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#148
post #30

This is awesome. No preview release either, which is great to production. They are pushing the prices higher with each release though: API pricing is up to $0.5/M for input and $3/M for output For comparison: Gemini 3.0 Flash: $0.50/M for input and $3.00/M for output Gemini 2.5 Flash: $0.30/M for input and $2.50/M for output Gemini 2.0 Flash: $0.15/M for input and $0.60/M for output Gemini 1.5 Flash: $0.075/M for inp…

is there a website where i can compare openai, anthropic and gemini models on cost/token ?

Re: Gemini 3 Flash: Frontier intelligence built for speed

#149
post #125

At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…

That's a quite sensationalized view.

Ghibli moment was only about half a year ago. At that moment, OpenAI was so far ahead in terms of image editing. Now it's behind for a few months and "it can't be reversed"?

Re: Gemini 3 Flash: Frontier intelligence built for speed

#150
post #125

At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…

OAI's latest image model outperforms Google's in LMArena in both image generation and image editing. So even though some people may prefer nano banana pro in their own anecdotal tests, the average person prefers GPT image 1.5 in blind evaluations.

https://lmarena.ai/leaderboard/text-to-image

https://lmarena.ai/leaderboard/image-edit

Post reply on HN