Interesting to note that this might be the only model with knowledge cut off as recent as 2025 January
Gemini 2.5 Flash
21–30 of 582 posts
Re: Gemini 2.5 Flash
#22This is cool, but rate limits on all of these preview models are PITA
Re: Gemini 2.5 Flash
#23Absolutely decimated on metrics by o4-mini, straight out of the gate, and not even that much cheaper on output tokens (o4-mini's thinking can't be turned off IIRC).
Re: Gemini 2.5 Flash
#24Gemini flash models have the least hype, but in my experience in production have the best bang for the buck and multimodal tooling. Google is silently winning the AI race.
Re: Gemini 2.5 Flash
#25OpenAI might win the college students but it looks like Google will lock in enterprise.
I built a product that uses and LLM and I got curious about the quality of the output from different models. It took me a weekend to go from just using OpenAI's API to having Gemini, Claude, and DeepSeek all as options and a lot of that time was research on what model from each provider that I wanted to use.
Re: Gemini 2.5 Flash
#26Interesting to note that this might be the only model with knowledge cut off as recent as 2025 January
Re: Gemini 2.5 Flash
#27Gemini flash models have the least hype, but in my experience in production have the best bang for the buck and multimodal tooling. Google is silently winning the AI race.
accuracy | input price | output price
Gemini Flash 2.0 Lite: 67% | $0.075 | $0.30
Gemini Flash 2.0: 93% | $0.10 | $0.40
GPT-4.1-mini: 93% | $0.40 | $1.60
GPT-4.1-nano: 43% | $0.10 | $0.40
excited to to try out 2.5 flash
Re: Gemini 2.5 Flash
#28OpenAI might win the college students but it looks like Google will lock in enterprise.
Funny you should say that. Google just announced today that they are giving all college students one year of free Gemini advanced. I wonder how much that will actually move the needle among the youth.
Re: Gemini 2.5 Flash
#29Bad day is going on google. First the decleration of illegal monopoly.. and now... Google’s latest innovation: programmable overthinking. With Gemini 2.5 Flash, you too can now set a thinking_budget—because nothing says "state-of-the-art AI" like manually capping how long it’s allowed to reason. Truly the dream: debugging a production outage at 2am wondering if your LLM didn’t answer correctly because you cheaped out…
Re: Gemini 2.5 Flash
#30Gemini flash models have the least hype, but in my experience in production have the best bang for the buck and multimodal tooling. Google is silently winning the AI race.