Live data from Hacker News

Gemini 2.5 Flash

developers.googleblog.com

551–560 of 582 posts

Re: Gemini 2.5 Flash

#551
post #205

Earlier quoted context omitted.

More and more people are coming to the realisation that Google is actually winning at the model level right now.

I haven’t met a single person that uses Gemini. Companies are using Copilot and individuals are using ChatGPT. Also, why would I want Google to spy on my AI usage? They’re evil.

why is Google more evil than say OpenAI ?

Re: Gemini 2.5 Flash

#552

One hidden note from Gemini 2.5 Flash when diving deep into the documentation: for image inputs, not only can the model be instructed to generated 2D bounding boxes of relevant subjects, but it can also create segmentation masks! https://ai.google.dev/gemini-api/docs/image-understanding#se... At this price point with the Flash model, creating segmentation masks is pretty nifty. The segmentation masks are a bit of a g…

There is a starter app in AI Studio that demos this: https://aistudio.google.com/apps/bundled/spatial-understandi...

Re: Gemini 2.5 Flash

#553
post #548
post #545

Earlier quoted context omitted.

Generative Pre-trained Transformer is a horrible term to have an acronym for.

Do you think the mass market thinks GPT is an acronym? It's just a name. Currently synonymous with AI. Ask anyone outside the tech bubble about "Gemini" though. You'll get astrology.

True I guess they treat it just like SMS.

I still think they'd have taken off more if they'd given it a catchy name from the start and made the interface a bit more consumer friendly.

Re: Gemini 2.5 Flash

#554
post #207
post #192

Earlier quoted context omitted.

After comparing Gemini Pro and Claude Sonnet 3.7 coding answers side by side a few times, I decided to cancel my Anthropic subscription and just stick to Gemini.

Google has killed so many amazing businesses -- entire industries, even, by giving people something expensive for free until the competition dies, and then they enshittify hard. It's cool to have access to it, but please be careful not to mistake corporate loss leaders for authentic products.

How would I know if it’s useful to me without being able to trial it?

Googles previous approach (Pro models available only to Gemini Advanced subscribers, and Advanced trials can’t be stacked with Google One paid storage, or rather they convert the already paid storage portion to a paid, much shorter Advanced subscription!) was mind-bogglingly stupid.

Having a free tier on all models is the reasonable option here.

Re: Gemini 2.5 Flash

#555

Earlier quoted context omitted.

Though the same logic can be applied to everywhere, right? Even if it's done by human interns, you need to audit everything to be 100% confident or just have some trust on them.

Not the same logic because interns can make meaning out of the data - that’s built-in error correction. They also remember what they did - if you spot one misunderstanding, there’s a chance they’ll be able to check all similar scenarios. Comparing the mechanics of an LLM to human intelligence shows deep misunderstanding of one, the other, or both - if done in good faith of course.

Not sure why you're trying to conflate intellectual capability problems into this and complicate the argument? The problem layout is the same. You delegate the works to someone so you cannot understand all the details. This makes a fundamental tension between trust and confidence. Their parameters might be different due to intellectual capability, but whoever you're going to delegate, you cannot evade this trade-off.

BTW, not sure if you have experiences of delegating some works to human interns or new grads and being rewarded by disastrous results? I've done that multiple times and don't trust anyone too much. This is why we typically develop review processes, guardrails etc etc.

Re: Gemini 2.5 Flash

#557
post #205

Earlier quoted context omitted.

More and more people are coming to the realisation that Google is actually winning at the model level right now.

What’s with the Google cheer squad in this thread, usually it’s Google lost its way and is evil. Can’t be employees cause usually there is a disclaimer

Gemini 2.5 is genuinely impressive.

Re: Gemini 2.5 Flash

#558

Earlier quoted context omitted.

This might have changed after you posted your comment, but it looks like 2.5 Pro and 2.5 Flash are available in the Gemini app now, both web and mobile.

Oh, I didn’t mean to say that these models were unavailable through the app or website. Rather, I’ve realized that using them through the API or AI Studio yields much better results — even in the free tier. You can check that by trying prompts with complex instructions and long inputs/outputs. For instance, ask Gemini to generate notes from a specific source (say, a book or class transcription). Or ask it to translat…

Underutilized, or over-prompted for the layperson?

Re: Gemini 2.5 Flash

#559
post #379
post #62

Earlier quoted context omitted.

done pretty much inline with the price elo pareto frontier https://x.com/swyx/status/1912959140743586206/photo/1

So if I see it right flash 2.5 doesn't push the pareto front forward, right? It just sits between 2.5 pro and 2.0 flash. https://storage.googleapis.com/gweb-developer-goog-blog-asse...

It does, that point in the tradeoff space was not available until now. Any model that's not dominated by at least one model on both axes will push forward the frontier. (The actual frontier isn't actually a straight line between the points on the frontier like visualized there. It's a step function.)

Re: Gemini 2.5 Flash

#560
post #528

Honestly, the best part about Gemini, especially as a consumer product, is their super lax, or lack thereof, ratelimits. They never have capacity issues, unlike Claude which always feels slow or sometimes outright rejects requests during peak hours. Gemini is constantly speedy and has extremely generous context window limits on the Gemini apps.

Interesting. I use Claude quite a bit, and haven't encountered this. Is this the free version of Claude or the paid version? When are peak hours typically (in what timezone)?

I have Claude Pro and peak hours are in the afternoon and at night for me in EST
Post reply on HN