Earlier quoted context omitted.
I think their downfall will be the fact that they don't have a "path to AGI" and have been raising investor money on the promise that they do.
I believethere’s also exponential dislike growing for Altman among most AI users, and that impacts how the brand/company is perceived.
Gemini 3 Flash: Frontier intelligence built for speed
511–520 of 609 posts
Re: Gemini 3 Flash: Frontier intelligence built for speed
#512Earlier quoted context omitted.
This has been my dream for voice control of PC for ages now. No wake word, no button press, no beeping or nagging, just fluently describe what you want to happen and it does.
without a wake word, it would have to listen and process all parsed audio. you really want everything captured near the device/mic to be sent to external servers?
Re: Gemini 3 Flash: Frontier intelligence built for speed
#513Re: Gemini 3 Flash: Frontier intelligence built for speed
#514This model is breaking records on my benchmark of choice, which is 'the fraction of Hacker News comments that are positive.' Even people who avoid Google products on principle are impressed. Hardly anyone is arguing that ChatGPT is better in any respect (except brand recognition).
No offense, but that seems like a poor benchmark. These initial vibe checks are easily swayed by personal brand biases.
I do pay special attention to what the most negative comments say (which in this case are unusually positive). And people discussing performance on their own personal benchmarks.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#515Earlier quoted context omitted.
Until recently, Google was the underdog in the LLM race and OpenAI was the reigning champion. How quickly perceptions shift!
I just want a deepseek moment for an open weights model fast enough to use in my app, I hate paying the big guys.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#516So gemini 3 flash (non thinking) is now the first model to get 50% on my "count the dog legs" image test. Gemini 3 pro got 20%, and everyone else has gotten 0%. I saw benchmarks showing 3 flash is almost trading blows with 3 pro, so I decided to try it. Basically it is an image showing a dog with 5 legs, an extra one photoshopped onto it's torso. Every models counts 4, and gemini 3 pro, while also counting 4, said th…
Re: Gemini 3 Flash: Frontier intelligence built for speed
#517Earlier quoted context omitted.
is that faster to say than do, or is it an accessibility or while-driving need?
I don't understand that use case at all. How can you tell it to do all that stuff, if you aren't sitting there glued to the screen yourself?
Plus, if the above worked, the higher level interactions could trivially work too. "Go to event details", "add that to my calendar".
FWIW, I'm starting to embrace using Gemini as general-purpose UI for some scenarios just because it's faster. Most common one, " add to my calendar please."
Re: Gemini 3 Flash: Frontier intelligence built for speed
#518Re: Gemini 3 Flash: Frontier intelligence built for speed
#519Earlier quoted context omitted.
Enterprise is slow. As for developers, we will be switching to Google unless the competition can catch up and deliver a similarly fast model. Enterprise will follow. I don't see any distinction in target markets - it's the same market.
Yeah, this is what I was trying to say in my original comment too. Also I do not really use agentic tasks but I am not sure that gemini 3/3 flash have mcp support/skills support for agentic tasks if not, I feel like they are very low hanging fruits and something that google can try to do too to win the market of agentic tasks over claude too perhaps.
So far they seem faster with Flash, and with less corruption of files using the Edit tool - or at least it recovered faster.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#520Earlier quoted context omitted.
Oh wow - I recently tried 3 Pro preview and it was too slow for me. After reading your comment I ran my product benchmark against 2.5 flash, 2.5 pro and 3.0 flash. The results are better AND the response times have stayed the same. What an insane gain - especially considering the price compared to 2.5 Pro. I'm about to get much better results for 1/3rd of the price. Not sure what magic Google did here, but would love…
Curious to learn what a “product benchmark” looks like. Is it evals you use to test prompts/models? A third party tool? Examples from the wild are a great learning tool, anything you’re able to share is appreciated.
And it shouldn't be shared publicly so that the models won't learn about it accidentally :)