Live data from Hacker News

Please don't discontinue Gemini 2.5 Flash

discuss.ai.google.dev

51–60 of 93 posts

Re: Please don't discontinue Gemini 2.5 Flash

#51

Earlier quoted context omitted.

Where do people get ideas like this? In what world does this make sense? You have several choices: 1. Work with a supplier and sign a contract guaranteeing support for whatever period of time you want at a mutually agreeable price 2. Host your own stack to depend on and support it for however long you want 3. Accept that you're paying for a service and that it can go away at any time. Companies aren't obligated to su…

It's not a legal obligation, no, but neither should you as a customer accept a vendor that treats you like that.

> vendor that treats you like that

You don't have to use Big Ai offerings, there are other options. Between deprecation and uncle sam, dependency/business risk appears to be increasing.

It's a calculation and choice that comes with consequences any way you land.

Re: Please don't discontinue Gemini 2.5 Flash

#53

I am more concerned about the cost step up from Gemini 2.5 Flash to 3.5 Flash, with the latter being roughly 3x more expensive. I thought the intention of the Flash models was to be relatively low-latency and more affordable compared to Pro, but the newer Flash models aren’t being priced as such. Then again, the era of cheap and plentiful AI might be coming to an end…

Yeah IIRC the latest Pro is $12 and Flash is $9 which is not the usual 2X-3X multiplier we see separate model grades. It also puts Flash now about 2X GLM 5.2, which is a highly capable open weight model.

I think the thing is that 3.5 flash is actually similarly capable on a lot of tasks that matter and is faster. Pro is more specialised in the direction of mathematical reasoning and stuff.

Re: Please don't discontinue Gemini 2.5 Flash

#54

Earlier quoted context omitted.

There are plenty of open weights models available already. If the ability to keep running the same model is important to you, then choose one of those.

Presumably, a “Stop Killing AI” movement, mirroring the Stop Killing Games movement, would require a provider that revokes access a previously available model to make it open weights at the time of death.

On the surface, there appears a difference between buying a game and paying for llm processing time. You haven't bought the model, so it is unclear to me why the same argument ought to hold up.

Re: Please don't discontinue Gemini 2.5 Flash

#55

sucks we use 2.5 flash/lite in our company it handles millions of requests a day theres nothing in its price range that provides the same all around perf as noted, gemini 3 flash is expensive really not liking google these days they are not hungry anymore

They appear to be trying lock-in, or some sort of way to make Gemini family the only logical choice on their cloud. They don't offer the most desired open weight models per-token, so we found another vendor and are less likely to use Google services going forward (for more reasons than this)

Re: Please don't discontinue Gemini 2.5 Flash

#57

Earlier quoted context omitted.

Is it really? A 9B model is equivalent? Honest question, as I haven't spent that much time with the 9B variant or Flash 2.5. But that seems like a pretty bold claim for such a small model. I assumed Flash 2.5 was considerably larger, but maybe I'm wrong?

Pretty much. It even beats it in a few benchmarks: https://artificialanalysis.ai/models/comparisons/qwen3-5-9b-... Qwen 3.6 and Gemma 4 small models are in a league of their own.

Second this, they keep getting performance upgrades too. Z Lab had been publishing dflash addons and boosting their tgen 2-3x. I'm looking at doing comparative evals right now

https://huggingface.co/collections/z-lab/dflash

Post reply on HN