Live data from Hacker News

Gemini 3 Flash: Frontier intelligence built for speed

blog.google

531–540 of 609 posts

Re: Gemini 3 Flash: Frontier intelligence built for speed

#531

Earlier quoted context omitted.

If Opus is one-size-fits-all, then why Claude keeps the other series? (rethorical). Opus and Sonnet are slower than Haiku. For lots of less sophisticated tasks, you benefit from the speed. All vendors do this. You need smaller models that you can rapid-fire for lots of other reasons than vibe coding. Personally, I actually use more smaller models than the sophisticated ones. Lots of small automations.

Yes, all the major CLIs (Claude Code, Codex, etc) and many agentic applications use a large model main agent with task delegation to small model sub-agent. For example in CC using Opus4.5 it will delegate an Explore task to a Haiku/Sonnet subagent or multiple subagents.

The agent interfaces are for human interaction. Some tasks can be fully unattended though. For those, I find smaller models more capable due to their speed.

Think beyond interfaces. I'm talking about rapid-firing hundreds of small agents and having zero human interaction with them. The feedback is deterministic (non agentic) and automated too.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#532
post #524
post #436

Earlier quoted context omitted.

Whether Googling something counts as AI has more to do with the shifting definition of AI over time, then with Googling itself. Remember, really back in the day the A* search algorithm was part of AI. If you had asked anyone in the 1970s whether a box that given a query pinpoints the right document that answers that question (aka Google search in the early 2000s), they'd definitely would have called it AI.

Google gives you an AI summary, reading that means interacting with LLMs.

Google also gives you ads. Some learn to scroll past before reading.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#533

I don't want to say OpenAI is toast for general chat AI, but it sure looks like they are toast.

I’ve fully switched over to Gemini now. It seems significantly more useful, and is less of an automatic glaze machine that just restates your question and how smart you are for asking it.

That’s funny, I’ve had the exact opposite experience. Gemini starts every answer to a coding question with, “you have hit upon a fundamental insight in zyx”. ChatGPT usually starts with, “the short answer? Xyz.”

Re: Gemini 3 Flash: Frontier intelligence built for speed

#534

Earlier quoted context omitted.

Apple’s most interesting value proposition was ignoring all this AI junk and letting users click “not interested” on Apple Intelligence and never see it again. From a business perspective it’s a smart move (inasmuch as “integrating AI” is the default which I fundamentally disagree with) since Apple won’t be left holding the bag on a bunch of AI datacenters when/if the AI bubble pops. I don’t want to lose trust in App…

> Apple’s most interesting value proposition was ignoring all this AI junk Did you forget all the Apple Intelligence stuff? They were never "ignoring" if anything they talked a big talk, and then failed so hard. The whole iPhone 16 was marketed as AI first phone (including in billboards). They had full length ads running touting AI benefits. Apple was never "ignoring" or "sitting AI out". They were very much in it. A…

[deleted]

Re: Gemini 3 Flash: Frontier intelligence built for speed

#535
post #124

Earlier quoted context omitted.

Really stupid question: How is Gemini-like 'thinking' separate from artificial general intelligence (AGI)? When I ask Gemini 3 Flash this question, the answer is vague but agency comes up a lot. Gemini thinking is always triggered by a query. This seems like a higher-level programming issue to me. Turn it into a loop. Keep the context. Those two things make it costly for sure. But does it make it an AGI? Surely Googl…

Advanced reasoning LLM's simulate many parts of AGI and feel really smart, but fall short in many critical ways. - An AGI wouldn't hallucinate, it would be consistent, reliable and aware of its own limitations - An AGI wouldn't need extensive re-training, human reinforced training, model updates. It would be capable of true self-learning / self-training in real time. - An AGI would demonstrate real genuine understand…

> - It should even demonstrate consciousness.

I disagreed with most of your assertions even before I hit the last point. This is just about the most extreme thing you could ask for. I think very few AI researchers would agree with this definition of AGI.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#536
post #363
post #352

Earlier quoted context omitted.

So this is an interesting benchmark, because if the answer is actually in the top 3 google results, then my python script that runs a google search, scrapes the top n results and shoves them into a crappy LLM would pass your benchmark too! Which also implies that (for most tasks), most of the weights in a LLM are unnecessary, since they are spent on memorizing the long tail of Common Crawl... but maybe memorizing inf…

I've tried doing this query with search enabled in LLMs before, which is supposed to effectively do that, and even with that they didn't give very good answers. It's a very physical kind of thing, and its easy to conflate with other similar descriptions, so they would frequently just conflate various different things and give some horrible mash-up answer that wasn't about the specific thing I'd asked about.

So it's a difficult question for LLMs to answer even when given perfect context?

Kinda sounds like you're testing two things at the same time then, right? The knowledge of the thing (was it in the training data and was it memorized?) and the understanding of the thing (can they explain it properly even if you give them the answer in context).

Re: Gemini 3 Flash: Frontier intelligence built for speed

#537
post #344

Earlier quoted context omitted.

I desperately want to be able to real-time dictate actions to take on my phone. Stuff like: "Open Chrome, new tab, search for xyz, scroll down, third result, copy the second paragraph, open whatsapp, hit back button, open group chat with friends, paste what we copied and send, send a follow-up laughing tears emoji, go back to chrome and close out that tab" All while being able to just quickly glance at my phone. Ther…

This has been my dream for voice control of PC for ages now. No wake word, no button press, no beeping or nagging, just fluently describe what you want to happen and it does.

Apple tried this ages ago:

https://en.wikipedia.org/wiki/PlainTalk

Re: Gemini 3 Flash: Frontier intelligence built for speed

#538
post #4

Don’t let the “flash” name fool you, this is an amazing model. I have been playing with it for the past few weeks, it’s genuinely my new favorite; it’s so fast and it has such a vast world knowledge that it’s more performant than Claude Opus 4.5 or GPT 5.2 extra high, for a fraction (basically order of magnitude less!!) of the inference time and price

Lately I was trying ask LLMs to generate SVG pictures, do you have famous pelican on bike created by flash model?

Re: Gemini 3 Flash: Frontier intelligence built for speed

#539

Earlier quoted context omitted.

I wonder at what point will everyone who over-invested in OpenAI will regret their decision (expect maybe Nvidia?). Maybe Microsoft doesn't need to care, they get to sell their models via Azure.

But you’re forgetting the Jonny Ive hardware device that totally isn’t like that laughable pin badge thing from Humane /s

I agree completely. Altman was at some point talking about a screen less device and getting people away from the screen.

Abandoning our mose useful sense, vision, is a recipe for a flop.

Re: Gemini 3 Flash: Frontier intelligence built for speed

#540
post #273
post #238

Feels like Google is really pulling ahead of the pack here. A model that is cheap, fast and good, combined with Android and gsuite integration seems like such powerful combination. Presumably a big motivation for them is to be first to get something good and cheap enough they can serve to every Android device, ahead of whatever the OpenAI/Jony Ive hardware project will be, and way ahead of Apple Intelligence. Speakin…

What will you use the ai in the phone to do for you? I can understand tablets and smart glasses being able to leverage smol AI much better than a phone which is reliant on apps for most of the work.

Analyse e-mails/text/music/videos, edit photos, summarization, etc.
Post reply on HN