Live data from Hacker News

Gemini 3.7 Flash

blog.google

61–70 of 525 posts

Re: Gemini 3.7 Flash

#63

They compare it to 5.6 Terra, however https://cognition.com/frontiercode puts Terra at about 1/2 the price Also have to compare to the recent Grok 4.6 release, which appears to straight up be better AND cheaper Hard to understand why anyone would choose 3.7 Flash under these conditions.. is Deepmind still a frontier lab?

Agree, with you but I'm still using 3.6 Flash because of tok/s/ latency/ uptime with high context. Tried Grok 4.6 and it was scoring lower on some internal benchmarks or slower.

Re: Gemini 3.7 Flash

#65
Did the company fix the high friction between any service and their models' API?

I hope so. It seems mind boggling to me that an user needs to surf around different sections (plural) of google cloud console, then this Vertex and do a dozen clicks to issue a simple key.

Re: Gemini 3.7 Flash

#66
Grok, Meta, Gemini and others all released updates to their models within around a month or two from their respective last release and made significant jumps in benchmarks all around the same time. Any guesses as to why that is? Is it just the release season and/or everyone is benchmaxxing?

Re: Gemini 3.7 Flash

#67

Does Google believe people want fast models because they have some sort of evidence of that preference? Or are they no longer capable of delivering a Pro model?

If you think about Google and their business/reach, fast and light models suite them the best.

Google probably crunches more tokens daily than the other labs combined, just because basically the entire global population uses Google (sans china) and Google has shoved Gemini into everything.

Re: Gemini 3.7 Flash

#68

They compare it to 5.6 Terra, however https://cognition.com/frontiercode puts Terra at about 1/2 the price Also have to compare to the recent Grok 4.6 release, which appears to straight up be better AND cheaper Hard to understand why anyone would choose 3.7 Flash under these conditions.. is Deepmind still a frontier lab?

Artificial Analysis shows Grok 4.6 taking $1,068 to run their suite while Gemini 3.7 Flash takes $485. So it looks like Gemini 3.7 Flash is less than half the price in the real world.

Per-token cost isn't a great metric given that some use way more tokens than others.

Re: Gemini 3.7 Flash

#69

Does Google believe people want fast models because they have some sort of evidence of that preference? Or are they no longer capable of delivering a Pro model?

I have read that "pro"/"opus"/etc models can actually be worse for everyday coding as they reason "too deeply" and turn over too many stones over-thinking the problem and potentially getting distracted.

This feels absurd to me (my gut is "I want the SMARTEST model I can get!!"), but often I find that my experience of using a flash/sonnet model for every-day workhorse coding they are better.

Its not the same thing, but when I think of that I am reminded of working with some engineers in the past who are incredibly smart and have PhDs (or to put it another way, over-qualified) and they were crap engineers because they'd just not be able to focus on the task and ONLY the task at hand and would get easily distracted by the "why" or "more interesting" things when I just asked them to fix a simple bug or whatever. Again, its not the same thing at all, but it certainly comes to mind when I think of this or experience a pro/opus model suggesting we make huge refactors when a tactical fix is all that is required etc.

Of course, the opus-sized models are great when it comes to huge comprehension/research/debugging efforts where the deeper reasoning is actually useful.

Re: Gemini 3.7 Flash

#70
post #41

They need to release benchmarks against Luna/Terra. Luna is much cheaper which feels like it undercuts the need for Flash. I've always considered the Flash series of models to be for low-cost, high-volume, mostly text-based use cases (e.g. summarization, parsing, formatting), emphasis on low-cost. [edit: ah, benchmarks here: https://blog.google/innovation-and-ai/models-and-research/ge... more of a Terra than Luna com…

gemini flash is probably the best model for visual tasks right now. they also make it really easy to ingest videos
Post reply on HN