Live data from Hacker News

Gemini 3.7 Flash

blog.google

281–290 of 525 posts

Re: Gemini 3.7 Flash

#281

Earlier quoted context omitted.

Can you help me understand how it is hard to get an API key from Google? You just head on over to http://aistudio.google.com/api-keys and create a key... not any different from platform.openai.com? Disclaimer: I work in Google so it might be that this link is not publicly well known

On top of being the hardest website to navigate, Google console a) doesn't have real time billing (!) b) doesn't allow you to set a budget limit. Sorry but it's not worth waking up with a 100k$ bill, fix your platform first.

Oh it’s worse than that. There are places where you CAN set a limit. This seems great until you are at the center of a huge traffic spike because of good pr and so you try to change it to a larger number only to be told you need to wait 24 hours for the setting to change.

Biggest traffic day of the decade and our site was down because of google.

Re: Gemini 3.7 Flash

#282
post #193

Earlier quoted context omitted.

Sol on Cerebras is going to be expensive AF

Moving from either frontier intelligence or frontier latency to a single model that does both at the same time is potentially a game changer in certain industries. I can easily see e.g. hedge funds dropping tons of money on this, because it means they can now do the same thing as their competitors, but much faster. That's basically a license to print money.

What would a hedge fund want to do on this exactly? It’s too slow for hft and I’m not sure what they would be doing where ms matter but is not hft.

Re: Gemini 3.7 Flash

#283
post #60

https://artificialanalysis.ai/models/gemini-3-7-flash The selling point for gemini continues to be speed and particularly end-to-end response time.

I've blown away by flash 3.6's speed while Opus chugs along for _hours_ on similar tasks. I've gotten into a opus designed -> gemini implemented -> opus reviewed dev cycle recently.

> I've gotten into a opus designed -> gemini implemented -> opus reviewed dev cycle recently.

This is what I do too.

Re: Gemini 3.7 Flash

#285
post #224

Earlier quoted context omitted.

Can you help me understand how it is hard to get an API key from Google? You just head on over to http://aistudio.google.com/api-keys and create a key... not any different from platform.openai.com? Disclaimer: I work in Google so it might be that this link is not publicly well known

Disclaimer that I haven't tried this since January, so things may have changed in the last 7mo, but this was my experience at that time: https://x.com/pwnies/status/2010523020629274723 At a high level though, as a rule of thumb Google assumes that they're serving companies at Google scale first, and at a human scale second. For other companies it's the opposite. Generally what that means is the first experience you g…

Don't forget your account getting flagged for review, so you get to wait an extra 24 hours for no reason.

Re: Gemini 3.7 Flash

#286
On a related note, I see all these quantitative benchmarks and the models getting really good at them over time. One thing I've been wondering: if the GPT series of models performs so well quantitatively, why do I still kind of hate using them relative to Claude? There’s a missing “vibes” or “taste” benchmark I think.

Re: Gemini 3.7 Flash

#287
post #133
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

Other thoughts: I really think Google has fallen behind here. Even as a high speed offering (this build took ~7min, which is pretty good!), it wont be able to claim dominance for long with cerebras announcing the Sol preview today: https://www.cerebras.ai/blog/accelerating-gpt-5-6-sol-ultraf... . It's not a bad model by any means, but I just don't know what situation I'd reach for 3.7 Flash first for. Google really n…

> especially given how hard it is to get an API key from them

What does this mean? Anybody can get an API key

Re: Gemini 3.7 Flash

#288
post #141

The "introductory pricing" for this 3.7 Flash model is really weird. It's scheduled to double in price on December 31, 2026, but who would anticipate still using this model five months from now? Especially since 3.6 Flash came out just three weeks ago! My first effort with default thinking level produced an ambitious pelican, let down by a flawed bicycle: https://tools.simonwillison.net/markdown-svg-renderer#url=ht..…

I assume so when they release Flash 3.8 or 4.0 or whatever they can claim it's not a price increase?

Re: Gemini 3.7 Flash

#289
post #235
post #224

Earlier quoted context omitted.

Disclaimer that I haven't tried this since January, so things may have changed in the last 7mo, but this was my experience at that time: https://x.com/pwnies/status/2010523020629274723 At a high level though, as a rule of thumb Google assumes that they're serving companies at Google scale first, and at a human scale second. For other companies it's the opposite. Generally what that means is the first experience you g…

FWIW I definitely did not not need to do anything like that to generate a key via AI Studio. It was like three clicks to get the free tier key, later on enabling billing was a few more plus typing in credit card info.

yeah, "enabling billing" is a whole other ordeal. Like it all makes sense, it's not that hard, it's just a lot of extra clicks where the competition doesn't require that. You can blow off this feedback as the user whining, but it's real friction and the competition doesn't have that, so users are going to go elsewhere if they can.

Re: Gemini 3.7 Flash

#290
post #193

Earlier quoted context omitted.

Sol on Cerebras is going to be expensive AF

Moving from either frontier intelligence or frontier latency to a single model that does both at the same time is potentially a game changer in certain industries. I can easily see e.g. hedge funds dropping tons of money on this, because it means they can now do the same thing as their competitors, but much faster. That's basically a license to print money.

I am not sure this is the way to make AI more cost effective for such customers. If they are able to tweak any model for their use case it would be way more reliable and also way cheaper. In my opinion generic LLMs in the future will be just for attention economy or maybe government contracts. Everyone else will be running fine tuned free weight models or licenced closed source models (self hosted or managed).
Post reply on HN