Live data from Hacker News

Gemini 3.7 Flash

blog.google

301–310 of 525 posts

Re: Gemini 3.7 Flash

#301
Have you tried the new DeepSeek Pro v4, Qwen 3.8, Gemini 3.7 Flash, and Grok 4.6? Do they make sense for any use cases?

I'm currently using omp with Kimi K3 as the planner and DeepSeek v4 Flash 0731 as the implementer, or CC + Fable for planning and Opus 4.8 for implementation. For API(not coding), I just use DeepSeek v4 flash 0731 and MiMo.

I'm pretty happy where I am, but I'm wondering if these new models provide some new kind of advantage

Re: Gemini 3.7 Flash

#302
post #82

IMO, they should drop their previous model (3.6 Flash) from the benchmark charts. I don't care how better this is compared with their previous model. What matters (to me) is: 1. How the new model performs against the other top models in the same category. 2. The pricing of the new model against the other top models in the same category.

[deleted]

Re: Gemini 3.7 Flash

#303
post #294
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

Can you share what your prompt was for that?

The original prompt includes auth codes for API gen, but here's a redacted version: https://non.io/prompt-for-ramen

This is the build step generated by my diffusion-based ui tool's copy-for-agent action.

Re: Gemini 3.7 Flash

#304

Earlier quoted context omitted.

Moving from either frontier intelligence or frontier latency to a single model that does both at the same time is potentially a game changer in certain industries. I can easily see e.g. hedge funds dropping tons of money on this, because it means they can now do the same thing as their competitors, but much faster. That's basically a license to print money.

What would a hedge fund want to do on this exactly? It’s too slow for hft and I’m not sure what they would be doing where ms matter but is not hft.

There's a lot of trading that isn't proper "HFT", but where speed and latency still matter. Often you'll find this employed more as slippage reduction - i.e you're going to make the trade either way, but making it faster saves you a few bps.

I'm not sure what event-based traders are doing now, but back in the day NLP sentiment analysis was all the rage, so I'm assuming they've now incorporated LLMs too.

Re: Gemini 3.7 Flash

#305
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

This has so much less character than the pelican smdh.. Plus is the ramen in HK even any good?

I think they're both testing very different things. The pelican test is testing if a LLM can come up with visuals on its own via writing bezier curves directly.

This is testing if it can match visuals that have already been established, and represent them with all the tools available to a web developer. The ramen example was chosen in particular because there are a lot of things that aren't easy to do with CSS, and require creative strategies: 45deg button cuts, angular repeating pattern elements, blending of raster art and svgs, microglyphs, low contrast subtle elements, etc.

Don't ask yourself whether it's a good design, as yourself whether it's a good test.

Re: Gemini 3.7 Flash

#306
Google has got the IBM disease.

Large lumbering enterprise with massive inertia. Where innovators leave as soon as they get a better offer.

None of the authors of the seminal "Attention Is All You Need" paper are still at Google.

Fast forward a decade and Google will be reduced to hiring the kind of mediocrities who deign to work at IBM and Accenture.

Re: Gemini 3.7 Flash

#307
post #186
post #169

Earlier quoted context omitted.

I already use GCP and Google for work, and getting an API key was so annoying that even I couldn't be bothered after a while of looking around. Maybe things there have improved some, but when I was looking it was a huge runaround.

Hmm. My company has an internal portal for generating Gemini API keys. I select a project from a drop down, enter a name, and press okay.

That may be evidence the built-in Google experience is difficult or confusing.

Re: Gemini 3.7 Flash

#308
post #52
post #41

They need to release benchmarks against Luna/Terra. Luna is much cheaper which feels like it undercuts the need for Flash. I've always considered the Flash series of models to be for low-cost, high-volume, mostly text-based use cases (e.g. summarization, parsing, formatting), emphasis on low-cost. [edit: ah, benchmarks here: https://blog.google/innovation-and-ai/models-and-research/ge... more of a Terra than Luna com…

They compared against 5.6-terra on the model card: https://deepmind.google/models/model-cards/gemini-3-7-flash/

[deleted]

Re: Gemini 3.7 Flash

#309
post #240
post #70

Earlier quoted context omitted.

gemini flash is probably the best model for visual tasks right now. they also make it really easy to ingest videos

Also best at OpenSCAD, seemingly for the same reason, at least in terms of "iterate on a design, comparing visual output to target".

Are you manually rendering previews of its OpenSCAD output to create images for it to review, or do you have a workflow that automates that?

Re: Gemini 3.7 Flash

#310
post #106

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this. Original images: https://image.non.io/neonRamenDesigns.webp Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7 Opus 5 build for comparison: https://html.non.io/neonRamen Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM pr…

Was the original concept generated by Claude somehow? It gives me Claude UI vibes with all the extraneous small-caps text elements.
Post reply on HN