Live data from Hacker News

Gemini 3.7 Flash

blog.google

421–430 of 525 posts

Re: Gemini 3.7 Flash

#421
post #133
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

Other thoughts: I really think Google has fallen behind here. Even as a high speed offering (this build took ~7min, which is pretty good!), it wont be able to claim dominance for long with cerebras announcing the Sol preview today: https://www.cerebras.ai/blog/accelerating-gpt-5-6-sol-ultraf... . It's not a bad model by any means, but I just don't know what situation I'd reach for 3.7 Flash first for. Google really n…

The API key you are mentioning is just ridiculous. Onboarding your company or personal account is a trap. I ended up getting assigned to sales guy just to test their Vertex API because I used a company email.

Of course, we just used OpenRouter for testing and never touched a Gemini model anymore.

Re: Gemini 3.7 Flash

#422
post #394
post #391

Earlier quoted context omitted.

Sergey's out spending $100 million to lobby against the billionaire tax in California. Lest his net worth shrink from roughly $250 billion to $237 billion. Can't become a pleb now! [0] [0] https://techcrunch.com/2026/08/10/google-co-founder-sergey-b...

not that I'm intending to defend Sergey here but, if you could spend 0.04% of your net worth to protect 5% of your net worth (and probably all of your easy liquidity)... wouldn't you?

If you mean given my current net worth, then the question doesn't make sense.

If you mean a world where I had his net worth, then you're accusing me of being a sociopath.

Re: Gemini 3.7 Flash

#423
post #224

Earlier quoted context omitted.

Can you help me understand how it is hard to get an API key from Google? You just head on over to http://aistudio.google.com/api-keys and create a key... not any different from platform.openai.com? Disclaimer: I work in Google so it might be that this link is not publicly well known

Disclaimer that I haven't tried this since January, so things may have changed in the last 7mo, but this was my experience at that time: https://x.com/pwnies/status/2010523020629274723 At a high level though, as a rule of thumb Google assumes that they're serving companies at Google scale first, and at a human scale second. For other companies it's the opposite. Generally what that means is the first experience you g…

expertise at navigating accidental complexity can easily be mistaken for engineering expertise.

Re: Gemini 3.7 Flash

#424
post #133

Earlier quoted context omitted.

Other thoughts: I really think Google has fallen behind here. Even as a high speed offering (this build took ~7min, which is pretty good!), it wont be able to claim dominance for long with cerebras announcing the Sol preview today: https://www.cerebras.ai/blog/accelerating-gpt-5-6-sol-ultraf... . It's not a bad model by any means, but I just don't know what situation I'd reach for 3.7 Flash first for. Google really n…

Can you help me understand how it is hard to get an API key from Google? You just head on over to http://aistudio.google.com/api-keys and create a key... not any different from platform.openai.com? Disclaimer: I work in Google so it might be that this link is not publicly well known

The problem is that this doesn’t work for enterprise. The rate limits of that is super low. Then you need to migrate to Vertex and that is just a pain. Who ever thought of using JSON instead of an api key…

Re: Gemini 3.7 Flash

#425
post #394
post #391

Earlier quoted context omitted.

Sergey's out spending $100 million to lobby against the billionaire tax in California. Lest his net worth shrink from roughly $250 billion to $237 billion. Can't become a pleb now! [0] [0] https://techcrunch.com/2026/08/10/google-co-founder-sergey-b...

not that I'm intending to defend Sergey here but, if you could spend 0.04% of your net worth to protect 5% of your net worth (and probably all of your easy liquidity)... wouldn't you?

I think that argument should stop with advocacy for law. I have no problem with this when it's means playing optimally within the current laws, things like hiring world class tax attorneys. But if using your immense wealth to ensure the rules that everyone plays by favor you specifically, instead of the country/world, It feels like this is closer to bribing a judge than hiring a tax attorney.

Re: Gemini 3.7 Flash

#426
Funnily enough, I just ran a task on AI Studio with 3.6 yesterday and got 3.7 to do a similar one today; so it serves as an interesting and quick comparisons between the old and the new (usually, if enough time passes between your use of one model and the next, you'll have a sourer view of it than its actual competence suggests).

It hallucinated in both cases despite being given an API key and building a lot of pipes to access data using this. It was a simple "oh shit" fix moment for the model, but weird how eager it was to hallucinate despite the process being designed for it to be data-driven.

We should move past the idea that benchmarks alone tell us whether a model is getting better. I would've had the same experience a year or two ago with 1.5, and the solution would've been similar (keep prompting). I've been investing time into making system prompts and input prompts more meticulous, but the fundamental "it will make shit up" problem still remains, even though it shouldn't when the job involves calling tools.

I know this sounds like I'm expecting superpowers of it (I'm not), but my point is just that these incremental benchmark gains may not reflect user experience.

Re: Gemini 3.7 Flash

#427
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

I'm not sure what you consider good design, but if it's subjective, then I see it differently from your examples.

Gemini 3.7 looks the best. Opus 5 looks almost as good as Gemini. Grok 4.6 looks pretty terrible.

Re: Gemini 3.7 Flash

#428
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

IMO Gemini's is better than all the others, including the original.

Re: Gemini 3.7 Flash

#429
I couldn't get past the first chart which basically showed that intelligence and cost are both better for gpt luna, I'm not sure what the argument is here for Gemini flash 3.7 given that comparison.
Post reply on HN