Live data from Hacker News

Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

blog.google

451–460 of 616 posts

Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

#451
post #442

Here's the issue: GLM 5.2 is better, also cheaper, and almost as fast. So essentially, a big L for Google. Combine this with them not being able to produce a frontier model this generation... hmm implications

3.6 is roughly 50% faster, which isn't totally insignificant for being marginally more expensive.[1] [1]artificialanalysis.ai

Sure, but there's no sota alternative from Google. That's it, and its beaten by GLM 5.2 on every measure except somewhat speed.

I find that quite staggering. GLM is open weights

Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

#452
post #79
post #71

Pelicans for 3.6 Flash and 3.5 Flash-Lite (Cyber isn't available to me through the API yet.) https://tools.simonwillison.net/markdown-svg-renderer#url=ht...

I am growing tired of these pelicans posts every time a new model is published. Feels to me like low effort personal brand promotion. Just sharing my 2 cents.

Your comment reads very pedantic with a hint of jealousy. The pelican and xbox controllers are great ways to see how well it can follow direction dealing with svg a difficult format for LLMs to use and testing their spatial vision awareness.

Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

#453

Pricing per million input/output tokens: 2.5 Flash: $0.3 / $2.5 3.0 Flash: $0.5 / $3 3.5 Flash: $1.5 / $9 3.6 Flash: $1.5 / $7.5 --- 2.5 Flash-Lite: $0.1 / $0.4 3.1 Flash-Lite: $0.25 / $1.5 3.5 Flash-Lite: $0.3 / $2.5

2.5 flash was the only reason we were paying four digits a month to Google....

i guess we'll use 3.0 flash but thats going to get replaced too right ?

these flash lite models aren't very reliable or consistent

Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

#454
post #390

I often use Gemini free web chat because it's generally quite good at web search-related questions (apparently it has direct token-level access to the Google Search index) but I noticed in the last two weeks output quality of 3.5 Flash seriously degraded. Maybe they were switching over systems.

Google's Knowledge Graph is a massive advantage no other competitor has. I don't think they've fully utilized its full potential but I don't know if any other company could've built something like Scholar Labs

Knowledge Graph + automated transcriptions of almost every YouTube video = giant untapped moat of data

Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

#455
post #107

Earlier quoted context omitted.

It's also very possible that they know their big model underperforms chatgpt 5.6 and fable by too much, so they are focusing on what they can get wins in like speed instead.

> focusing on what they can get wins in like speed instead Speed as a differentiator has always been Google's thing. They (used to?) show the microseconds it took to query & rank web-scale search results. Chrome, notoriously, focused on speed at the expense of resource use. The very many efforts to efficiently speed up Android & its runtime since its inception, and so on... > their big model underperforms chatgpt 5.6…

That sonds like they can't compete with 3.5 or 3.6 so they must increase the model size and are training v4.

Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

#456
post #136

Always happy to see new Gemini releases as IMO Antigravity Pro 16.67/mo plan (Annual) is still the best plan available and have been pretty happy with Antigravity IDE. If it wasn't for Gemini/Antigravity I'd have to go with a Max Claude plan, as it stands now I can get by with just a Claude Pro plan to get Opus when I need it, whilst using Antigravity as my day-to-day workhorse. Unfortunately Gemini Flash became too…

What's the current outlook on Antigravity IDE vs. 2.0? How long will they begrudgingly keep it going before kicking everyone to 2.0/3.0?

(I actually use a mix of both for some offline projects, nothing serious.)

Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

#458

LLM reception is truly extreme, even worse than AAA game releases. Ever frontier lab lived it at least once : missing the frontier by a few months triggers extremly negative reactions, then you take back the lead for 2 weeks, and the hype cycle repeats.

I spin a mental roulette on whether the reception on a new release will be "OMG best model by far, no one will be able to catch up for months!" or "OMG this is already outdated, RIP company X, they might as well just give up now, there's no coming back from this."

It's quite a fun game. I click into the comments and see if the roulette wheel was right.

Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

#459

Earlier quoted context omitted.

> Domestic China is the only very large audience for their own models I don't think so. US models are very expensive, and not available in every country. I am not willing to pay $50/1M tokens for writing my pet projects.

There are also US based companies like Fireworks serving up the best open weight models with the compliances we need in US enterprise. Depending on the company, they may offer more/different jurisdictions, EU probably needs a Fireworks like company (haven't heard about one, maybe it already exists?)

There is at least doubleword.ai, and there should be others.

Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

#460
post #79

Earlier quoted context omitted.

I am growing tired of these pelicans posts every time a new model is published. Feels to me like low effort personal brand promotion. Just sharing my 2 cents.

I find Simon's work informative and entertaining; the last thing he can be accused of is low effort. The Pelicans are just a bit of fun icing on top.

Sending a one sentence prompt to an LLM and posting it to hackernews constantly isn’t low effort? Today I learnt something new
Post reply on HN