This is the first flash/mini model that doesn't make a complete ass of itself when I prompt for the following: "Tell me as much as possible about Skatval in Norway. Not general information. Only what is uniquely true for Skatval." Skatval is a small local area I live in, so I know when it's bullshitting. Usually, I get a long-winded answer that is PURE Barnum-statement, like "Skatval is a rural area known for its bea…
Gemini 3 Flash: Frontier intelligence built for speed
201–210 of 609 posts
Re: Gemini 3 Flash: Frontier intelligence built for speed
#202At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…
Google has incredible tech. The problem is and always has been their products. Not only are they generally designed to be anti-consumer, but they go out of their way to make it as hard as possible. The debacle with Antigravity exfiltrating data is just one of countless.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#203Earlier quoted context omitted.
Is there a "good enough" endgame for LLMs and AI where benchmarks stop mattering because end users don't notice or care? In such a scenario brand would matter more than the best tech, and OpenAI is way out in front in brand recognition.
For average consumers, I think very much yes, and this is where OpenAI's brand recognition shines. But for anyone using LLM's to help speed up academic literature reviews where every detail matters, or coding where every detail matters, or anything technical where every detail matters -- the differences very much matter. And benchmarks serve just to confirm your personal experience anyways, as the differences between…
We've seen this movie before. Snapchat was the darling. Infact, it invented the entire category and was dominating the format for years. Then it ran out of time.
Now very few people use Snapchat, and it has been reduced to a footnote in history.
If you think I'm exaggerating, that just proves my point.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#204At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…
I'm actually liking 5.2 in Codex. It's able to take my instructions, do a good job at planning out the implementation, and will ask me relevant questions around interactions and functionality. It also gives me more tokens than Claude for the same price. Now, I'm trying to white label something that I made in Figma so my use case is a lot different from the average person on this site, but so far it's my go to and I d…
It's when it becomes difficult, like in the coding case that you mentioned, that we can see the OpenAI still has the lead. The same is true for the image model, prompt adherence is significantly better than Nano Banana. Especially at more complex queries.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#205Earlier quoted context omitted.
For average consumers, I think very much yes, and this is where OpenAI's brand recognition shines. But for anyone using LLM's to help speed up academic literature reviews where every detail matters, or coding where every detail matters, or anything technical where every detail matters -- the differences very much matter. And benchmarks serve just to confirm your personal experience anyways, as the differences between…
> OpenAI's brand recognition shines. We've seen this movie before. Snapchat was the darling. Infact, it invented the entire category and was dominating the format for years. Then it ran out of time. Now very few people use Snapchat, and it has been reduced to a footnote in history. If you think I'm exaggerating, that just proves my point.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#206Earlier quoted context omitted.
Is there a "good enough" endgame for LLMs and AI where benchmarks stop mattering because end users don't notice or care? In such a scenario brand would matter more than the best tech, and OpenAI is way out in front in brand recognition.
For average consumers, I think very much yes, and this is where OpenAI's brand recognition shines. But for anyone using LLM's to help speed up academic literature reviews where every detail matters, or coding where every detail matters, or anything technical where every detail matters -- the differences very much matter. And benchmarks serve just to confirm your personal experience anyways, as the differences between…
In fact so far, they consistently fail in exactly these scenario, glossing over random important details whenever you double check results in depth.
You might have found models, prompts or workflows that work for you though, I'm interested.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#207Earlier quoted context omitted.
Is there a "good enough" endgame for LLMs and AI where benchmarks stop mattering because end users don't notice or care? In such a scenario brand would matter more than the best tech, and OpenAI is way out in front in brand recognition.
this. I don't know any non-tech people who use anything other than chatgpt. On a similar note, I've wondered why Amazon doesn't make a chatgpt-like app with their latest Alexa+ makeover, seems like a missed opportunity. The Alexa app has a feature to talk to the LLM in chat mode, but the overall app is geared towards managing devices.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#208At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…
Is there anything pointing to Brin having anything to do with Google’s turnaround in AI? I hear a lot of people saying this, but no one explaining why they do
Founders are special, because they are not beholden to this social support network to stay in power and founders have a mythos that socially supports their actions beyond their pure power position. The only others they are beholden too are their co-founders, and in some cases major investor groups. This gives them the ability to disregard this social balance because they are not dependent on it to stay on power. Their power source is external to the organization, while everyone else is internal to it.
This gives them a very special "do something" ability that nobody else has. It can lead to failures (zuck & occulus, snapchat spectacles) or successes (steve jobs, gemini AI), but either way, it allows them to actually "do something".
Re: Gemini 3 Flash: Frontier intelligence built for speed
#209Earlier quoted context omitted.
“And then imagine Google designing silicon that doesn’t trail the industry. While you are there we may as well start to imagine Google figures out how to support a product lifecycle that isn’t AdSense” Google is great on the data science alone, every thing else is an after thought
https://blog.google/products/google-cloud/ironwood-google-tp... "And then imagine Google designing silicon that doesn’t trail the industry." I'm def not a Google stan generally, but uh, have you even been paying attention? https://en.wikipedia.org/wiki/Tensor_Processing_Unit
TPUs on the other hand are ASICs, we are more than familiar with the limited application, high performance and high barriers to entry associated with them. TPUs will be worthless as the AI bubble keeps deflating and excess capacity is everywhere.
The people who don't have a rudimentary understanding are the wall street boosters that treat it like the primary threat to Nvidia or a moat for Google (hint: it is neither).
Re: Gemini 3 Flash: Frontier intelligence built for speed
#210At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…
Is there a "good enough" endgame for LLMs and AI where benchmarks stop mattering because end users don't notice or care? In such a scenario brand would matter more than the best tech, and OpenAI is way out in front in brand recognition.