I love how there is a "Please do not discontinue gemini-2.0-flash[-lite], 2.5 is NOT an equivalent" from Feb 20th. Getting too attached to models is a smell.
the 1.5 and 2.0 flash models were absolute beasts. They were very cheap, and _very_ fast. We contemplated moving some of our fine tuned workloads to them because we would have gotten very substantial total latency reductions for our workloads. However, they are aggressively deprecating them (OpenAI is as well), and replacing with newer models. These newer models are all reasoning models, and importantly, only bear th…
Please don't discontinue Gemini 2.5 Flash
61–70 of 93 posts
Re: Please don't discontinue Gemini 2.5 Flash
#62I don't know how many times people will need to learn this: Do not use Google in Production.
in this case, the other frontier shops are all doing the same thing, so maybe the advice here should be "don't use closed weight models in production"?
Re: Please don't discontinue Gemini 2.5 Flash
#63I suppose at least in this case the loss is not an emotional one?
Re: Please don't discontinue Gemini 2.5 Flash
#64Earlier quoted context omitted.
in this case, the other frontier shops are all doing the same thing, so maybe the advice here should be "don't use closed weight models in production"?
Google has a history of pulling the rug on their paying customers and offering 0 support when it is convenient for them. You have no recourse. There are a billion cloud providers to choose from, this is not the first time this has been on the first page of HN.
Re: Please don't discontinue Gemini 2.5 Flash
#65Why not a "stop killing AI" movement? If a company deploys a paid AI model and makes people depend on it, they need to dump the weights at EOL.
If you are paying an ongoing subscription for a service, I'd advise you not to rely on it too much or keep a list of alternatives.
Re: Please don't discontinue Gemini 2.5 Flash
#66I love how there is a "Please do not discontinue gemini-2.0-flash[-lite], 2.5 is NOT an equivalent" from Feb 20th. Getting too attached to models is a smell.
It's not a smell. Why should these developers rebuild a core piece of their stack every few months. Switching out a model requires a new round of testing and validation when we should be able to rely on a piece of software the behave the same way since the last time we touched it.
Re: Please don't discontinue Gemini 2.5 Flash
#67Earlier quoted context omitted.
There are plenty of open weights models available already. If the ability to keep running the same model is important to you, then choose one of those.
Presumably, a “Stop Killing AI” movement, mirroring the Stop Killing Games movement, would require a provider that revokes access a previously available model to make it open weights at the time of death.
If you were paying for an ongoing subscription for a service, that would be something different.
Re: Please don't discontinue Gemini 2.5 Flash
#68It's such a good model for the price, for a lot of tasks it outperforms gpt5 at 3x the speed and 1/5 the price. The price jump from 2.5->3->3.5 has been so high.
Re: Please don't discontinue Gemini 2.5 Flash
#69It's such a good model for the price, for a lot of tasks it outperforms gpt5 at 3x the speed and 1/5 the price. The price jump from 2.5->3->3.5 has been so high.
i have found google models outperforming other models in actual agentic workflows
Re: Please don't discontinue Gemini 2.5 Flash
#70I love how there is a "Please do not discontinue gemini-2.0-flash[-lite], 2.5 is NOT an equivalent" from Feb 20th. Getting too attached to models is a smell.
the 1.5 and 2.0 flash models were absolute beasts. They were very cheap, and _very_ fast. We contemplated moving some of our fine tuned workloads to them because we would have gotten very substantial total latency reductions for our workloads. However, they are aggressively deprecating them (OpenAI is as well), and replacing with newer models. These newer models are all reasoning models, and importantly, only bear th…