> Today, we are releasing updated versions of Gemini 2.5 Flash and 2.5 Flash-Lite, available on Google AI Studio and Vertex AI, aimed at continuing to deliver better quality while also improving the efficiency. Typo in the first sentence? "... improving the efficiency." Gemini 2.5 Pro says this is perfectly good phrasing, whereas ChatGPT and Claude recognize that it's awkward or just incorrect. Hmm...
Improved Gemini 2.5 Flash and Flash-Lite
91–100 of 285 posts
Re: Improved Gemini 2.5 Flash and Flash-Lite
#92Google seems to be the main foundation model provider that's really focusing on the latency/TPS/cost dimensions. Anthropic/OpenAI are really making strides in model intelligence, but underneath some critical threshold of performance, the really long thinking times make workflows feel a lot worse in collaboration-style tools, vs a much snappier but slightly less intelligent model. It's a delicate balance, because thes…
I would be surprised if this dichotomy you're painting holds up to scrutiny. My understanding is Gemini is not far behind on "intelligence", certainly not in a way that leaves obvious doubt over where they will be over the next iteration/model cycles, where I would expect them to at least continue closing the gap. I'd be curious if you have some benchmarks to share that suggest otherwise. Meanwhile, afaik something G…
Re: Improved Gemini 2.5 Flash and Flash-Lite
#93Seems llm progress really is plateauing. I guess that was to be expected.
I actually even agree that the progress is plateauing, but your comment is a non-sequitur.
Re: Improved Gemini 2.5 Flash and Flash-Lite
#94This really captures something I've been experiencing with Gemini lately. The models are genuinely capable when they work properly, but there's this persistent truncation issue that makes them unreliable in practice. I've been running into it consistently, responses that just stop mid-sentence, not because of token limits or content filters, but what appears to be a bug in how the model signals completion. It's been…
Re: Improved Gemini 2.5 Flash and Flash-Lite
#95Am I the only one who is starting to feel the Gemini Flash models are better than Pro? Flash is super fast, gets straight to the point. Pro takes ages to even respond, then starts yapping endlessly, usually confuses itself in the process and ends up with a wrong answer.
Also 2.5 Pro is often incapable of searching and will hallucinate instead. I don't know why. It will claim it searched and then return some made up results instead. 2.5 Flash is much more consistently capable of searching
Re: Improved Gemini 2.5 Flash and Flash-Lite
#96Earlier quoted context omitted.
Why is Grok so popular
Grok Code Fast 1 usage is driven almost entirely by Kilo Code and Cline: https://openrouter.ai/x-ai/grok-code-fast-1/apps Both apps have offered usage for free for a limited time: https://blog.kilocode.ai/p/grok-code-fast-get-this-frontier-... https://cline.bot/blog/grok-code-fast
If xAI in particular is in the mood to light cash on fire promoting their new model, you'll see it everywhere during the promo period, so not surprised that heavily boosts xAI stats. The mystery codename models of the week are a bit easier to miss.
Re: Improved Gemini 2.5 Flash and Flash-Lite
#97Non-AI Summary: Both models have improved intelligence on Artificial Analysis index with lower end-to-end response time. Also 24% to 50% improved output token efficiency (resulting in lower cost). Gemini 2.5 Flash-Lite improvements include better instruction following, reduced verbosity, stronger multimodal & translation capabilities. Gemini 2.5 Flash improvements include better agentic tool use and more token-effici…
2.5 Flash is the first time I've felt AI has become truly useful to me. I was #1 AI hater but now find myself going to the Gemini app instead of Google search. It's just better in every way and no ads. The info it provides is usually always right and it feels like I have the whole generalized and accurate knowledge of the internet at my fingertips in the app. It's more intimate, less distractions. Just me and the Gem…
Re: Improved Gemini 2.5 Flash and Flash-Lite
#98Seems llm progress really is plateauing. I guess that was to be expected.
Re: Improved Gemini 2.5 Flash and Flash-Lite
#99Earlier quoted context omitted.
chatgpt also has lots of reliability issues
If anyone from OpenAI is reading this, I have two complaints: 1. Using the "Projects" thing (Folder organization) makes my browser tab (on Firefox) become unusably slow after a while. I'm basically forced to use the default chats organization, even though I would like to organize my chats in folders. 2. After editing a message that you already sent,you get to select between the different branches of the chat (1/2, an…