Live data from Hacker News

Gemini 3.7 Flash

blog.google

511–520 of 525 posts

Re: Gemini 3.7 Flash

#511
post #335

Earlier quoted context omitted.

I practically switched to doing everything with Luna or DeepSeek V4 flash. I haven't feel the need for the more expensive models.

Of the two, which do you find better?

I prefer V4 flash, but Luna is ok and included in the OpenAI plan I'm already paying, so... that is the reason I use it.

I gave DeepSeek $50 around june, and I haven't been able to exhaust them yet. The model is super cheap and more than enough for my needs.

I'm my opinion, Flash V4 is less pedantic than OpenAI models, less prone to unsolicited prescriptions and less prone to "helpfully" reinterpreting my instructions (wrongly, of course).

Re: Gemini 3.7 Flash

#512
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

How are you prompting it with the original images to create such sites?

Re: Gemini 3.7 Flash

#513
post #417
post #399

Earlier quoted context omitted.

For comparison tho, 400GB of cloud storage for $5/mo is actually quite amazing even if it didn’t include anything else. I know, it’s cloud storage, a NAS lets you own your data, but for non techies who have a bunch of photos and videos; it seems like an easy recommendation.

Is it though? Hetzner gives you 1 TB of cloud storage for 4 USD per month.

[deleted]

Re: Gemini 3.7 Flash

#514

Earlier quoted context omitted.

Quantitatively, yes. Qualitatively, no.

Qualitatively I’m much closer to a homeless person in terms of my power to control other people. We should stop calling it a wealth tax and start calling it an unearned power tax. That’s a far more apt term.

I meant in real practical terms, not higher concepts.

I am fine with not having a private island or a network of influential friends, but I am really really glad I’m not defecating in the street.

Re: Gemini 3.7 Flash

#515
post #447

Earlier quoted context omitted.

Modern protocols loop back the reasoning tokens in raw textual form via an encrypted parameter. You can't see them (modulo the recent attack), but you do resubmit them.

Yeah I've done more research and that's what I meant in the edit you replied to. But that's not the full reasoning token context, just a snapshot of the latent state at the end of it, no? Have a look at the gemini ones they're pretty small.

There was an attack that convinced models to leak the contents of the reasoning tokens, and it came back as text that matched the length of the encrypted data very well. So it's probably still tokens.

Looping back latent state isn't that easy. The hidden data is the entire contents of the KV cache which can be massive. I don't think any provider is trying to loop the KV cache through the client, and neural compressions of the reasoning would be lossy / an advanced technique that is still firmly in the realm of research papers, as I understand.

Re: Gemini 3.7 Flash

#516

They compare it to 5.6 Terra, however https://cognition.com/frontiercode puts Terra at about 1/2 the price Also have to compare to the recent Grok 4.6 release, which appears to straight up be better AND cheaper Hard to understand why anyone would choose 3.7 Flash under these conditions.. is Deepmind still a frontier lab?

Gemini models are still good for knowledge as per omniscience benchmarks on artificial analysis

Re: Gemini 3.7 Flash

#518

Earlier quoted context omitted.

Moving from either frontier intelligence or frontier latency to a single model that does both at the same time is potentially a game changer in certain industries. I can easily see e.g. hedge funds dropping tons of money on this, because it means they can now do the same thing as their competitors, but much faster. That's basically a license to print money.

I am not sure this is the way to make AI more cost effective for such customers. If they are able to tweak any model for their use case it would be way more reliable and also way cheaper. In my opinion generic LLMs in the future will be just for attention economy or maybe government contracts. Everyone else will be running fine tuned free weight models or licenced closed source models (self hosted or managed).

This was a widespread opinion a few years ago, but by now it is pretty much accepted that any LLM fine tune you build today based on the best available models will be beaten by a general purpose frontier LLM within a year.

Re: Gemini 3.7 Flash

#519

Earlier quoted context omitted.

Exactly that. Except that ultrafast delivers this level of intelligence an order of magnitude faster. So if your competitors automatically react to news articles or financial statements with a certain level of comprehension within minutes, you can now do so in seconds.

Except they all do, and the money flows to Cerebras :)

Eventually they will. But until these models become commonly available there is a significant first mover advantage.

Re: Gemini 3.7 Flash

#520
post #507

Earlier quoted context omitted.

Crazy it's still the only video understanding endpoint. It's what I use it for and no other model even offers a competitor.

to be fair, all it's doing is sampling the frames and maybe doing transcription, if I'm not mistaken. So you can do it with the other models too, you just need to sample the frames yourself and do the transcript yourself

it does this at a variable rate of frames which you can set - not sure if it is transcribing or natively understanding audio, but I think it's the latter since it is much faster than most transcription models I am aware of

Regardless you are right - I can roll my own.... but why

Post reply on HN