Gemini 3.7 Flash
401–410 of 525 posts
Re: Gemini 3.7 Flash
#402Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…
One of my favorite image tests with AI models is schematic analysis...I build and repair tube amps for a living, and use AI for such work a LOT. so far, IMHO, the best has been opus and fable\mythos.
Re: Gemini 3.7 Flash
#403Earlier quoted context omitted.
flash-lite is more of their luna tier competitor but even still not quite there yet, but gemini's dominance on multimodal and image understanding i think really gets downplayed on this site when most people think the only think you can do with LLMs is write code
Ultimately it would track that in the real world, people will want to point cameras at things and get answers. I pay for ChatGPT and Gemini, and while Sol is a total beast with anything text, it still poisoned my cucumber bed. Which I will be bitter about for at least a few years while the bed recovers. Gemini (even flash) is exceptionally talented at viewing photos and telling you what to do/what it is (and telling…
Re: Gemini 3.7 Flash
#404Earlier quoted context omitted.
> I've gotten into a opus designed -> gemini implemented -> opus reviewed dev cycle recently. This is what I do too.
Last time I tried it, Flash introduced too many errors due to sloppiness. Is it more reliable at following instructions now?
Re: Gemini 3.7 Flash
#405Earlier quoted context omitted.
Sergey's out spending $100 million to lobby against the billionaire tax in California. Lest his net worth shrink from roughly $250 billion to $237 billion. Can't become a pleb now! [0] [0] https://techcrunch.com/2026/08/10/google-co-founder-sergey-b...
not that I'm intending to defend Sergey here but, if you could spend 0.04% of your net worth to protect 5% of your net worth (and probably all of your easy liquidity)... wouldn't you?
Folks need to remember that we're closer to being homeless than we are to being as rich as them. You don't need to defend billionaires.
Re: Gemini 3.7 Flash
#406The "introductory pricing" for this 3.7 Flash model is really weird. It's scheduled to double in price on December 31, 2026, but who would anticipate still using this model five months from now? Especially since 3.6 Flash came out just three weeks ago! My first effort with default thinking level produced an ambitious pelican, let down by a flawed bicycle: https://tools.simonwillison.net/markdown-svg-renderer#url=ht..…
maybe exactly because the frontier moves, introductory pricing makes sense as you want to free up compute for the newer models
Re: Gemini 3.7 Flash
#407Recently my android phone updated from Google assistant to Gemini Flash. Completely unusable. Asking it to play music and it refuses, hallucinating instructions to connect Spotify to Gemini. The instructions say to tap buttons that don't exist. Bonus feature from Gemini: a toggle to opt back into Google Assistant, but it doesn't work. Still stuck with Gemini.
Re: Gemini 3.7 Flash
#408Earlier quoted context omitted.
5.6 Luna costs far less and benchmarks far better, have you compared for this task?
Uh snap it indeed is cheaper: $0.20 / $1.20 vs $0.30 / $2.50. Gemini is mostly good enough for what I do with it, but the cost savings are interesting. Gemini is still faster though. Not sure how much the benchmarks can be trusted though: https://www.reddit.com/r/GoogleGeminiAI/comments/1vbq5vf/com...
Re: Gemini 3.7 Flash
#409Ever since the insane discount with GPT-5.6 Luna, not much excites me anymore. I mean just look at the benchmarks, even though Gemini 3.7 Flash performs well on the DeepSWE 1.1, Luna (Max) still performs way better. I personally have stuck to Luna (Xhigh) because its been more than enough and does not bloat up the context window too fast with reasoning tokens. https://deepswe.datacurve.ai > Starting January 1, 2027,…
Re: Gemini 3.7 Flash
#410Earlier quoted context omitted.
Disclaimer that I haven't tried this since January, so things may have changed in the last 7mo, but this was my experience at that time: https://x.com/pwnies/status/2010523020629274723 At a high level though, as a rule of thumb Google assumes that they're serving companies at Google scale first, and at a human scale second. For other companies it's the opposite. Generally what that means is the first experience you g…
Similar experience here, for what it's worth, though I didn't get as far as you. I basically just stopped and didn't bother - it was easier to go through OpenRouter than spend more energy on it. Also: > Google assumes that they're serving companies at Google scale first So much this. I'm currently grandfathered in until the end of the year on Google's Search API, but the $35,000 they want to continue usage of my < 10…