Live data from Hacker News

Gemini 3.7 Flash

blog.google

91–100 of 525 posts

Re: Gemini 3.7 Flash

#91
post #9

The multimodal abilities are great, but if you deal with text only, what is the benefit of using this over DS V4 Flash/Pro? 13-26x cheaper with comparable intelligence, and available across many different inference providers. I fail to see the usecase where DS V4 Pro is not enough, but Flash 3.7 is - except multimodal. Luna is similar, and also 8x cheaper. Source: artificialanalysis The only benefit I can see is the…

From my own testing, Gemini 3.5/3.6 Flash is better than DS v4 Flash/Pro on text ability.

Re: Gemini 3.7 Flash

#92
post #21

I'm only interested in the state-of-the-art model by each provider. For Google, this is still gemini-3.1-pro-preview, right?

This is all a naming quirk because Google can’t commit Path A: Deprecated, do not dare use Path B: Beta, do not rely

Google once again seems to have fallen into the pit of its own bureaucracy, even OpenAI looks competent by comparison.

Re: Gemini 3.7 Flash

#93
post #12

> * For 3.6 and 3.7 Flash, introductory price expires on December 31, 2026. Starting January 1, 2027, $1.50/1M input tokens and $7.50/1M output tokens will apply. this is hilarious. it is not 2025 any more, by Jan 2027 there will be at least 3 newer generation of models (from other provider) released already. nobody would use flash 3.7 at that time. sure we used to cling to gemini models in the past, demanding 2.5 mo…

What? Jan 2027 is just about four months away. People surely still use models from four months ago today.

Re: Gemini 3.7 Flash

#94
post #12

> * For 3.6 and 3.7 Flash, introductory price expires on December 31, 2026. Starting January 1, 2027, $1.50/1M input tokens and $7.50/1M output tokens will apply. this is hilarious. it is not 2025 any more, by Jan 2027 there will be at least 3 newer generation of models (from other provider) released already. nobody would use flash 3.7 at that time. sure we used to cling to gemini models in the past, demanding 2.5 mo…

Nobody except corporations who built workflows on top of it and don't care about the price because the developer already moved on and nobody wants to touch it.

Re: Gemini 3.7 Flash

#95
post #66

Grok, Meta, Gemini and others all released updates to their models within around a month or two from their respective last release and made significant jumps in benchmarks all around the same time. Any guesses as to why that is? Is it just the release season and/or everyone is benchmaxxing?

its essentially the same model being trained continuously 24/7 with the company periodically publishing just a new checkpoint

each new checkpoint can benefit from better reasoning training, RL on specific tasks and more synthetic data

So why do they seem to release around the same time ? my guess is because they time major releases around quarterly earnings, investor meetings and other important business milestones. Once one company announces a major update, the others also have an incentive to ship their latest checkpoint rather than look like they r falling behind.

Re: Gemini 3.7 Flash

#97
post #41

They need to release benchmarks against Luna/Terra. Luna is much cheaper which feels like it undercuts the need for Flash. I've always considered the Flash series of models to be for low-cost, high-volume, mostly text-based use cases (e.g. summarization, parsing, formatting), emphasis on low-cost. [edit: ah, benchmarks here: https://blog.google/innovation-and-ai/models-and-research/ge... more of a Terra than Luna com…

flash-lite is more of their luna tier competitor but even still not quite there yet, but gemini's dominance on multimodal and image understanding i think really gets downplayed on this site when most people think the only think you can do with LLMs is write code

Re: Gemini 3.7 Flash

#98

Does Google believe people want fast models because they have some sort of evidence of that preference? Or are they no longer capable of delivering a Pro model?

All the leaks say their latest attempt at a Pro model was not competitive.

Especially not competitive at software engineering.

Re: Gemini 3.7 Flash

#99
post #83
post #12

> * For 3.6 and 3.7 Flash, introductory price expires on December 31, 2026. Starting January 1, 2027, $1.50/1M input tokens and $7.50/1M output tokens will apply. this is hilarious. it is not 2025 any more, by Jan 2027 there will be at least 3 newer generation of models (from other provider) released already. nobody would use flash 3.7 at that time. sure we used to cling to gemini models in the past, demanding 2.5 mo…

Isn't it a good thing to know about price hikes in advance? If I were building a product around it, I would certainly care.

I think it's meant to make fun of the fact that Google raised prices on their models and people were upset, and this is Google's way of lowering back the price because by Jan 1st 2027, this model isn't going to be used since people will move on to the latest models.

Personally, I feel like Google blundered on their pricing because while I was using the free version of the Gemini harness, they took away most of the free limits and made people move over to their Anti-Gravity harness for no apparent reason. I was about to splurge for a Pro sub since I already used Google for extra storage but putting up limits like they did made me not want to trust they wouldn't do more price shenanigans. Now their models are behind and it seems like they're scrambling.

Re: Gemini 3.7 Flash

#100
post #9

The multimodal abilities are great, but if you deal with text only, what is the benefit of using this over DS V4 Flash/Pro? 13-26x cheaper with comparable intelligence, and available across many different inference providers. I fail to see the usecase where DS V4 Pro is not enough, but Flash 3.7 is - except multimodal. Luna is similar, and also 8x cheaper. Source: artificialanalysis The only benefit I can see is the…

for non-coding applications, i think speed is a real differentiator. Im building an app that uses LLMs for some functionality that the user would not have any reason to expect is using AI and therefore having then wait seconds or minutes is just not feasible. latency is a huge upside for me
Post reply on HN