Live data from Hacker News

Gemini 3.7 Flash

blog.google

271–280 of 525 posts

Re: Gemini 3.7 Flash

#271
post #261
post #253

Earlier quoted context omitted.

I mean the thinking does not bloat the context window because it gets dropped at the next request.

It doesn't. It's called preserved reasoning and every recent reasoning model does it

Sorry, I've realised I was only partially correct.

Gemini[0] for example passes along a snapshot of the reasoning state but it's not the equivalent to keeping all the reasoning tokens in the context.

[0] https://ai.google.dev/gemini-api/docs/thinking#signatures

Edit: Apparently it does take the same space in the LLM latent space so I was wrong.

Re: Gemini 3.7 Flash

#272
post #235
post #224

Earlier quoted context omitted.

Disclaimer that I haven't tried this since January, so things may have changed in the last 7mo, but this was my experience at that time: https://x.com/pwnies/status/2010523020629274723 At a high level though, as a rule of thumb Google assumes that they're serving companies at Google scale first, and at a human scale second. For other companies it's the opposite. Generally what that means is the first experience you g…

FWIW I definitely did not not need to do anything like that to generate a key via AI Studio. It was like three clicks to get the free tier key, later on enabling billing was a few more plus typing in credit card info.

It was 3 clicks because you knew where to look for.

Re: Gemini 3.7 Flash

#273

For almost every section in the model card there is the message: Gemini 3.7 Flash is based on Gemini 3.6 Flash. Same training dataset, same software, same hardware, same architecture... I'm wondering what they changed actually for the model to be more powerful if the benchmark results are real and relevant. Maybe just tweak settings or the reasoning prompts and called it a new version of their model?

I work at GDM and this is not at all true, 3.7 is markedly different (and better) than 3.6.

Re: Gemini 3.7 Flash

#274
post #242
post #228

Earlier quoted context omitted.

You need to create a Google Cloud project to create an api key and when you try to create one you very often get error messages like: “Failed to create project, The request is suspicious. Please try again” or “ You do not have permission to create a key in this project”. You can then navigate multiple screens in GCP to make it work but it’s a hassle compared to any other provider (OAI/Ant/OpenRouter or any of the Chi…

I didn't have that issue back when I originally created my API keys a couple years ago. Just out of curiosity I switched to a different Google account that had never interacted with AI Studio and never used Google Cloud console. It was literally two clicks, and didn't even leave the page: the dialog asked to create a project and type in a name, I did that, clicked submit and then it was selected as the default projec…

Last time I used the AI studio free tier it was limited to one or two requests, effectively useless. I think people are complaining its too hard to set up a paid API key (no reason you should make paying customers spend more than a few clicks and a minute of their time to pay you)

Re: Gemini 3.7 Flash

#275
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

[dead]

Re: Gemini 3.7 Flash

#276
post #133

Earlier quoted context omitted.

Other thoughts: I really think Google has fallen behind here. Even as a high speed offering (this build took ~7min, which is pretty good!), it wont be able to claim dominance for long with cerebras announcing the Sol preview today: https://www.cerebras.ai/blog/accelerating-gpt-5-6-sol-ultraf... . It's not a bad model by any means, but I just don't know what situation I'd reach for 3.7 Flash first for. Google really n…

Can you help me understand how it is hard to get an API key from Google? You just head on over to http://aistudio.google.com/api-keys and create a key... not any different from platform.openai.com? Disclaimer: I work in Google so it might be that this link is not publicly well known

We have probably 10+ years old account with Google cloud etc. We recently had a production deployment, I went over to AI studio to get new keys and it kept failing saying "Failed to generate API key, The request is suspicious. Please try again" - It was through my standard browser, same geo-ip. And it just worked after 2 days.

Re: Gemini 3.7 Flash

#277

So at this point new models seem to only care about one task, software development. This really was not the original pitch of ai and I do not see how it justifies the insane spend or valuations it has produced.

Think of the potential layoffs of highly paid employees! But, I think it’s also based on what they are being used for, most LLM users are still mainly SWEs or similar as I understand and there’s a ton of data to train them for coding.

I agree its what they are being used for and their primary revenue source.

My point is mainly that was never the pitch that got ai the hype it did and imo doesn't justify the valuations even if we all lose our jobs to ai. Because it no longer seems like they even think its making other jobs go away.

Re: Gemini 3.7 Flash

#278

Earlier quoted context omitted.

Can you help me understand how it is hard to get an API key from Google? You just head on over to http://aistudio.google.com/api-keys and create a key... not any different from platform.openai.com? Disclaimer: I work in Google so it might be that this link is not publicly well known

> Can you help me understand how it is hard to get an API key from Google? Using Google products in general is an effing nightmare as soon as you have to give them money. The one thing you want in a business is to remove friction when people want to give you money, a concept Google has never been able to understand.

> Using Google products in general is an effing nightmare as soon as you have to give them money

Spending money via Google Pay on Android is extremely easy, Google does know how to accept customer's money (in the consumer space)

Re: Gemini 3.7 Flash

#279

I don't get it, Google could heavily subsidy their Gemini models to make it more attractive, but they prefer to not do it. I don't know one soul who is using Gemini models to code. Even OpenAI who doesn't have money or capacity is offering their Luna model at $1.2 per 1M/out.

Google's open secret is that they are also capturing spend on OAI and Anthropic tokens.

Re: Gemini 3.7 Flash

#280
post #141

The "introductory pricing" for this 3.7 Flash model is really weird. It's scheduled to double in price on December 31, 2026, but who would anticipate still using this model five months from now? Especially since 3.6 Flash came out just three weeks ago! My first effort with default thinking level produced an ambitious pelican, let down by a flawed bicycle: https://tools.simonwillison.net/markdown-svg-renderer#url=ht..…

I think introductory here means more or less permanent but they can't publicly admit there are no takers at a higher price. Anthropic for instance announced a couple of days ago that they are making Sonnet's 'introductory pricing' permanent https://xcancel.com/claudeai/status/2086891169217122586

When a company gives away service a heavily subsidized service as a promo, the full cost of serving it (compute) can get classified as sales and marketing instead of just cost of revenue, which makes your gross margin look better!
Post reply on HN