Live data from Hacker News

Gemini 3 Flash: Frontier intelligence built for speed

blog.google

111–120 of 609 posts

Re: Gemini 3 Flash: Frontier intelligence built for speed

#113

Disappointed to see continued increased pricing for 3 Flash (up from $0.30/$2.50 to $0.50/$3.00 for 1M input/output tokens). I'm more excited to see 3 Flash Lite. Gemini 2.5 Flash Lite needs a lot more steering than regular 2.5 Flash, but it is a very capable model and combined with the 50% batch mode discount it is CHEAP ($0.05/$0.20).

Have you seen any indications that there will be a Lite version?

I guess if they want to eventually deprecate the 2.5 family they will need to provide a substitute. And there are huge demands for cheap models.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#114

So much for "Monopolies get lazy, they just rent seek and don't innovate"

Monopolies and wanna-be monopolies on the AI-train are running for their lives. They have to innovate to be the last one standing (or second last) - in their mind.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#116
post #91
post #45

Earlier quoted context omitted.

Why wouldn't you switch? The cost to switch is near zero for me. Some tools have built in model selectors. Direct CLI/IDE plug-ins practically the same UI.

Not OP, but I feel the same way. Cost is just one of the factor. I'm used to Claude Code UX, my CLAUDE.md works well with my workflow too. Unless there's any significant improvement, changing to new models every few months is going to hurt me more.

I used to think this way. But I moved to AGENTS.md. Now I use the different UI as a mental context separation. Codex is working on Feature A, Gemini on feature B, Claude on Feature C. It has become a feature.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#117
post #28

Yet again Flash receives a notable price hike: from $0.3/$2.5 for 2.5 Flash to $0.5/$3 (+66.7% input, +20% output) for 3 Flash. Also, as a reminder, 2 Flash used to be $0.1/$0.4.

Yes, but this Flash is a lot more powerful - beating Gemini 3 Pro on some benchmarks (and pretty close on others). I don't view this as a "new Flash" but as "a much cheaper Gemini 3 Pro/GPT-5.2"

Right, depends on your use cases. I was looking forward to the model as an upgrade to 2.5 Flash, but when you're processing hundreds of millions of tokens a day (not hard to do if you're dealing in documents or emails with a few users), the economics fall apart.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#118
post #105

Earlier quoted context omitted.

Did you try with the grounding tool? Turning it on solved this problem for me.

what if the lie is a logical deduction error not a fact retrieval error

The error rate would still be improved overall and might make it a viable tool for the price depending on the usecase.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#120
post #30

This is awesome. No preview release either, which is great to production. They are pushing the prices higher with each release though: API pricing is up to $0.5/M for input and $3/M for output For comparison: Gemini 3.0 Flash: $0.50/M for input and $3.00/M for output Gemini 2.5 Flash: $0.30/M for input and $2.50/M for output Gemini 2.0 Flash: $0.15/M for input and $0.60/M for output Gemini 1.5 Flash: $0.075/M for inp…

For comparison, GPT-5 mini is $0.25/M for input and $2.00/M for output, so double the price for input and 50% higher for output.
Post reply on HN