Live data from Hacker News

Gemini 3 Flash: Frontier intelligence built for speed

blog.google

101–110 of 609 posts

Re: Gemini 3 Flash: Frontier intelligence built for speed

#103
post #75
post #30

This is awesome. No preview release either, which is great to production. They are pushing the prices higher with each release though: API pricing is up to $0.5/M for input and $3/M for output For comparison: Gemini 3.0 Flash: $0.50/M for input and $3.00/M for output Gemini 2.5 Flash: $0.30/M for input and $2.50/M for output Gemini 2.0 Flash: $0.15/M for input and $0.60/M for output Gemini 1.5 Flash: $0.075/M for inp…

Are these the current prices or the prices at the time the models were released?

Mostly at the time of release except for 1.5 Flash which got a price drop in Aug 2024.

Google has been discontinuing older models after several months of transition period so I would expect the same for the 2.5 models. But that process only starts when the release version of 3 models is out (pro and flash are in preview right now).

Re: Gemini 3 Flash: Frontier intelligence built for speed

#105
post #4

Don’t let the “flash” name fool you, this is an amazing model. I have been playing with it for the past few weeks, it’s genuinely my new favorite; it’s so fast and it has such a vast world knowledge that it’s more performant than Claude Opus 4.5 or GPT 5.2 extra high, for a fraction (basically order of magnitude less!!) of the inference time and price

Cool! I've been using 2.5 flash and it is pretty bad. 1 out of 5 answers it gives will be a lie. Hopefully 3 is better

Did you try with the grounding tool? Turning it on solved this problem for me.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#106
post #57

Ok, I was a bit addicted to Opus 4.5 and was starting to feel like there's nothing like it. Turns out Gemini 3 Flash is pretty close. The Gemini CLI is not as good but the model more than makes up for it. The weird part is Gemini 3 Pro is nowhere as good an experience. Maybe because its just so slow.

I will have to try that. Cursor bill got pretty high with Opus 4.5. Never considered opus before the 4.5 price drop but now it's hard to change... :)

Re: Gemini 3 Flash: Frontier intelligence built for speed

#107
post #68

Deepmind Page: https://deepmind.google/models/gemini/flash/ Developer Blog: https://blog.google/technology/developers/build-with-gemini-... Model Card [pdf]: https://deepmind.google/models/model-cards/gemini-3-flash/ Gemini 3 Flash in Search AI mode: https://blog.google/products/search/google-ai-mode-update-ge...

For anyone from the Gemini team reading this: these links should all be prominent in the announcement posts. I always have to hunt around for them!

Google actually does something similar for major releases - they publish a dedicated collection page with all related links.

For example, the Gemini 3 Pro collection: https://blog.google/products/gemini/gemini-3-collection/

But having everything linked at the bottom of the announcement post itself would be really great too!

Re: Gemini 3 Flash: Frontier intelligence built for speed

#108
post #105

Earlier quoted context omitted.

Cool! I've been using 2.5 flash and it is pretty bad. 1 out of 5 answers it gives will be a lie. Hopefully 3 is better

Did you try with the grounding tool? Turning it on solved this problem for me.

what if the lie is a logical deduction error not a fact retrieval error

Re: Gemini 3 Flash: Frontier intelligence built for speed

#109

It has a SimpleQA score of 69%, a benchmark that tests knowledge on extremely niche facts, that's actually ridiculously high (Gemini 2.5 *Pro* had 55%) and reflects either training on the test set or some sort of cracked way to pack a ton of parametric knowledge into a Flash Model. I'm speculating but Google might have figured out some training magic trick to balance out the information storage in model capacity. Tha…

This will be fantastic for voice. I presume Apple will use it

Re: Gemini 3 Flash: Frontier intelligence built for speed

#110
I tried Gemini CLI the other day, typed in two one line requests, then it responded that it would not go further because I ran out of tokens. I've hard other people complaint that it will re-write your entire codebase from scratch and you should make backups before even starting any code-based work with the Gemini CLI. I understand they are trying to compete against Claude Code, but this is not ready for prime time IMHO.
Post reply on HN