Live data from Hacker News

Gemini 3.7 Flash

blog.google

401–410 of 525 posts

Re: Gemini 3.7 Flash

#402
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

One of my favorite image tests with AI models is schematic analysis...I build and repair tube amps for a living, and use AI for such work a LOT. so far, IMHO, the best has been opus and fable\mythos.

Nice! I test if models know which tubes I can use for an amp given the power and number of pins.

Re: Gemini 3.7 Flash

#403

Earlier quoted context omitted.

flash-lite is more of their luna tier competitor but even still not quite there yet, but gemini's dominance on multimodal and image understanding i think really gets downplayed on this site when most people think the only think you can do with LLMs is write code

Ultimately it would track that in the real world, people will want to point cameras at things and get answers. I pay for ChatGPT and Gemini, and while Sol is a total beast with anything text, it still poisoned my cucumber bed. Which I will be bitter about for at least a few years while the bed recovers. Gemini (even flash) is exceptionally talented at viewing photos and telling you what to do/what it is (and telling…

Tell me more about your cucumber bed. We started some raised veggie beds this year and my wife is relying very heavily in ChatGPT and Claude for advice on how to deal with issues.

Re: Gemini 3.7 Flash

#404

Earlier quoted context omitted.

> I've gotten into a opus designed -> gemini implemented -> opus reviewed dev cycle recently. This is what I do too.

Last time I tried it, Flash introduced too many errors due to sloppiness. Is it more reliable at following instructions now?

It seems reliable to me now.

Re: Gemini 3.7 Flash

#405
post #394
post #391

Earlier quoted context omitted.

Sergey's out spending $100 million to lobby against the billionaire tax in California. Lest his net worth shrink from roughly $250 billion to $237 billion. Can't become a pleb now! [0] [0] https://techcrunch.com/2026/08/10/google-co-founder-sergey-b...

not that I'm intending to defend Sergey here but, if you could spend 0.04% of your net worth to protect 5% of your net worth (and probably all of your easy liquidity)... wouldn't you?

If my net worth were that high? No.

Folks need to remember that we're closer to being homeless than we are to being as rich as them. You don't need to defend billionaires.

Re: Gemini 3.7 Flash

#406
post #219
post #141

The "introductory pricing" for this 3.7 Flash model is really weird. It's scheduled to double in price on December 31, 2026, but who would anticipate still using this model five months from now? Especially since 3.6 Flash came out just three weeks ago! My first effort with default thinking level produced an ambitious pelican, let down by a flawed bicycle: https://tools.simonwillison.net/markdown-svg-renderer#url=ht..…

maybe exactly because the frontier moves, introductory pricing makes sense as you want to free up compute for the newer models

They wouldn’t be on the frontier at all if it weren’t for this pricing.

Re: Gemini 3.7 Flash

#407
post #395

Recently my android phone updated from Google assistant to Gemini Flash. Completely unusable. Asking it to play music and it refuses, hallucinating instructions to connect Spotify to Gemini. The instructions say to tap buttons that don't exist. Bonus feature from Gemini: a toggle to opt back into Google Assistant, but it doesn't work. Still stuck with Gemini.

It's weird because Gemini via Antigravity or Google Search "AI Mode" works great. It's just the Google Assistant replacement application that's terrible. I've completely given up trying to ask it anything via Android auto. I punch my directions manually in Google maps before I leave now because the voice assistant is unusable.

Re: Gemini 3.7 Flash

#408
post #221

Earlier quoted context omitted.

5.6 Luna costs far less and benchmarks far better, have you compared for this task?

Uh snap it indeed is cheaper: $0.20 / $1.20 vs $0.30 / $2.50. Gemini is mostly good enough for what I do with it, but the cost savings are interesting. Gemini is still faster though. Not sure how much the benchmarks can be trusted though: https://www.reddit.com/r/GoogleGeminiAI/comments/1vbq5vf/com...

You can use fast mode for 2.5x faster and it’ll still be cheaper on output tokens

Re: Gemini 3.7 Flash

#409

Ever since the insane discount with GPT-5.6 Luna, not much excites me anymore. I mean just look at the benchmarks, even though Gemini 3.7 Flash performs well on the DeepSWE 1.1, Luna (Max) still performs way better. I personally have stuck to Luna (Xhigh) because its been more than enough and does not bloat up the context window too fast with reasoning tokens. https://deepswe.datacurve.ai > Starting January 1, 2027,…

[deleted]

Re: Gemini 3.7 Flash

#410
post #224

Earlier quoted context omitted.

Disclaimer that I haven't tried this since January, so things may have changed in the last 7mo, but this was my experience at that time: https://x.com/pwnies/status/2010523020629274723 At a high level though, as a rule of thumb Google assumes that they're serving companies at Google scale first, and at a human scale second. For other companies it's the opposite. Generally what that means is the first experience you g…

Similar experience here, for what it's worth, though I didn't get as far as you. I basically just stopped and didn't bother - it was easier to go through OpenRouter than spend more energy on it. Also: > Google assumes that they're serving companies at Google scale first So much this. I'm currently grandfathered in until the end of the year on Google's Search API, but the $35,000 they want to continue usage of my < 10…

The current move to force the usage of their AI api from postpay to prepay is also another example. (At least it’s going to give me the last incentive to move to cheaper api)
Post reply on HN