Live data from Hacker News

Gemini 3.6 Flash

console.cloud.google.com

41–50 of 78 posts

Re: Gemini 3.6 Flash

#41
post #29

This model is not for builders and engineers. DeepSWE score of 49% is behind gpt 5.4 and muse spark. It's clearly intended to be an efficient model for google gemini usage. What is interesting is how this is announced before any Gemini Pro progress. From the outside it seems as though Google cannot keep up with other frontier models.

Who tested it on DeepSWE?

Edit: Oh it's in the other link

https://blog.google/innovation-and-ai/models-and-research/ge...

Re: Gemini 3.6 Flash

#42

Earlier quoted context omitted.

Antigravity is such a dumb name. An-ti-gra-vi-ty (5) is long, Ge-mi-ni (3) was better. Aren’t Marketing/PR people at FAANG supposed to be among the best in the industry?

Just as ChatGPT is now called "Chat" by the kids (go ahead and finish grinding your teeth, I'll wait...), if Antigravity takes off then it'll get an irksome nickname like Antigrav or just Grav.

how about gravy? or aunti-gravy?

Re: Gemini 3.6 Flash

#43
I'm not at all an industry pundit. But I suspect there's a reason we're not seeing leading models from Google recently.

Judging from my own frustrating attempts to use Gemini for vibe-coding, it seems like Google is badly over-sold (i.e., under-provisioned).

From all those promos giving away their pro-level subscription with phones; spinning up a mid-level subscription to undercut other providers and (probably most significantly) putting AI queries into ever search response because their flagship search product had become useless; they're promising a lot more processing to customers than they can reliably deliver.

The recent iterations seem to be intended not to push the capabilities forward, but to deliver capabilities at the current level while consuming less resources. That will allow them to maintain their trajectory until (I'm expecting) they get the huge infusion of extra compute resources from Space X later this year.

If I'm right, then I expect we should see Google start pushing forward again (rather than more of this lateral stuff) by the end of the year.

Re: Gemini 3.6 Flash

#44
post #34

It is both less intelligent and more expensive than GLM-5.2, while being closed weight.

Their niche is that you get the quality of a Chinese model for the price of an American model.

(Quoting myself from 2 days ago.)

Re: Gemini 3.6 Flash

#46

It's available in Antigravity as well - Gemini 3.6 Flash (High) Gemini 3.6 Flash (Medium) Gemini 3.6 Flash (Low) Gemini 3.5 Flash (Medium) Gemini 3.5 Flash (High) Gemini 3.5 Flash (Low) Gemini 3.1 Pro (Low) Gemini 3.1 Pro (High) Claude Sonnet 4.6 (Thinking) Claude Opus 4.6 (Thinking) GPT-OSS 120B (Medium)

Antigravity is such a dumb name. An-ti-gra-vi-ty (5) is long, Ge-mi-ni (3) was better. Aren’t Marketing/PR people at FAANG supposed to be among the best in the industry?

They actually paid [1] Lexicon Branding for this dumb ass name (probably millions of dollars).

[1] https://www.lexiconbranding.com/case-studies/google-antigrav...

Re: Gemini 3.6 Flash

#48
I would be fine lauding Gemini models if the only benefit of them was superior understanding of intent (read between the lines). I don't need it to code because other models are tuned for that explicitly, but I would like a model that is tuned to produce less mechanical output.

Re: Gemini 3.6 Flash

#49
post #37

I find their models are pretty good for answering every day questions, and with their large user base, I’m sure they’re getting lots of usage.

They also provide the most usage for free and in the cheap paid plans.

Re: Gemini 3.6 Flash

#50
post #31

It’s frankly embarrassing at this point. I’ve got free access through buying a Pixel phone and it’s not even worth using as it’s a waste of my time. Here’s my experience so far using it for basic sysadmin Linux type stuff. Gemini 3.1 Pro just feels a generation behind, from when models would miss easy things and make bad assumptions. Its not actively detrimental in bad way but the opportunity cost vs using something…

I respectfully disagree, at least for most tasks. 3.1 Pro: while it's coding performance is mediocre, a lot of coding work requires minimum actual thinking. I use it often for light refactoring, boilerplate generation, testcase skeleton generation, code review, language questions ("is there a better way to write this code block?"). I don't have a corporation behind me so costs matter. Considering that I'm a Pro subsc…

It was a decent deal ~6-8 months ago. I had been using 3.1 pro almost since release, but it really is feeling old. After using other models more in the past two months though...I really can't go back to 3.1 pro, as I just have to explain my reasoning so damn much to get it on the right path, where as opus or fable just "get it" from the context of the project much better.

Sonnet is roughly the same level as 3.1 pro for me.

Post reply on HN