Live data from Hacker News

Improved Gemini 2.5 Flash and Flash-Lite

developers.googleblog.com

181–190 of 285 posts

Re: Improved Gemini 2.5 Flash and Flash-Lite

#181

Serious question: If it's an improved 2.5 model, why don't they call it version 2.6? Seems annoying to have to remember if you're using the old 2.5 or the new 2.5. Kind of like when Apple released the third-gen iPad many years ago and simply called it the "new iPad" without a number.

It's pretty common to refer to models by the month and year they were released. For example, the latest Gemini 2.5 Flash is known as "google/gemini-2.5-flash-preview-09-2025" [1]. [1]: https://openrouter.ai/google/gemini-2.5-flash-preview-09-202...

If they're going to include the month and year as part of the version number, they should at least use big endian dates like gemini-2.5-flash-preview-2025-09 instead of 09-2025.

Re: Improved Gemini 2.5 Flash and Flash-Lite

#182
post #99

Earlier quoted context omitted.

It would also be nice if ChatGPT could move chats between projects. My sidebar is a nightmare.

You can drag and drop chats between projects

i know. i want the assistant to do it. shouldn't it be able to do work on its own platform?

Re: Improved Gemini 2.5 Flash and Flash-Lite

#183
Am I using a different Gemini from everyone else? We have Google Workspace at my job, so Gemini is baked in.

It is HORRENDOUS when compared to other models.

I hear a bunch of other people talking about how great Gemini is, but I've never seen it.

The responses are usually either incorrect, way too long, (essays when I wanted summaries) or just...not...good. I will ask the exact same question to both Gemini and ChatGPT (free) and GPT will give a great answer while the Gemini answer is trash.

Am I missing something?

Re: Improved Gemini 2.5 Flash and Flash-Lite

#184
I would really like to see the 270M but which also knows phonetic alphabetic pronounciation in sentences. Perhaps IPA?

I would like to try a small computer->human "upload" experiment, basic multilingual understanding without pronounciation knowledge would be very sad.

I intend to make a sort of computer reflexive game, I want to compare different upload strategies (with/without analog or classic error correcting codes, empirical spaced repetition constants, a ML predictor of which parameters I'm forgetting / losing resolution on.

Re: Improved Gemini 2.5 Flash and Flash-Lite

#185

Am I using a different Gemini from everyone else? We have Google Workspace at my job, so Gemini is baked in. It is HORRENDOUS when compared to other models. I hear a bunch of other people talking about how great Gemini is, but I've never seen it. The responses are usually either incorrect, way too long, (essays when I wanted summaries) or just...not...good. I will ask the exact same question to both Gemini and ChatGP…

Maybe you are using it wrong.

Re: Improved Gemini 2.5 Flash and Flash-Lite

#186

Am I using a different Gemini from everyone else? We have Google Workspace at my job, so Gemini is baked in. It is HORRENDOUS when compared to other models. I hear a bunch of other people talking about how great Gemini is, but I've never seen it. The responses are usually either incorrect, way too long, (essays when I wanted summaries) or just...not...good. I will ask the exact same question to both Gemini and ChatGP…

I've been finding it leaps and bounds above other models but I'm only using it via aistudio. I haven't tried any IDE integration or similar, so can't talk to that. I do still have to tell it to stop it with the effusive praise (I guess that also helps reduce context windows)

Re: Improved Gemini 2.5 Flash and Flash-Lite

#187
post #42

Earlier quoted context omitted.

Can't agree with that. Gemini doesn't lead just on price/performance - ironically it's the best "normie" model most of the time, despite it's lack of popularity with them until very recent. It's bad at agentic stuff, especially coding. Incomparably so compared to Claude and now GPT-5. But if it's just about asking it random stuff, and especially going on for very long in the same conversation - which non-tech users h…

My pet theory without any strong foundation is because OpenAI and Anthropic have trained their models really hard to fit the sycophantic mold of: =============================== Got it — *compliment on the info you've shared*, *informal summary of task*. *Another compliment*, but *downside of question*. ---------- (relevant emoji) Bla bla bla 1. Aspect 1 2. Aspect 2 ---------- *Actual answer* ----------- (checkmark e…

Gemini does the sycophantic thing too, so I'm not sure that holds water. I keep having to remind it to stop with the praise whenever my previous instruction slips out of context window.

Re: Improved Gemini 2.5 Flash and Flash-Lite

#188

Google seems to be the main foundation model provider that's really focusing on the latency/TPS/cost dimensions. Anthropic/OpenAI are really making strides in model intelligence, but underneath some critical threshold of performance, the really long thinking times make workflows feel a lot worse in collaboration-style tools, vs a much snappier but slightly less intelligent model. It's a delicate balance, because thes…

We had to drop Gemini api cause it was so unreliable in production, no matter how long you waited.

Re: Improved Gemini 2.5 Flash and Flash-Lite

#189
post #128

Why do all of these model providers have such issues naming/versioning them? Why even use a version number (2.5) if you aren't going to change it when you update the model? This industry desperately needs a Steve Jobs to bring some sanity to the marketing.

The version number is about the architecture of the model, the date is just about the last weights of the model.

we solved this problem like 30 years ago, just have a minor release, and you can always get the latest minor release
Post reply on HN