Live data from Hacker News

Qwen 3.8

twitter.com

141–150 of 793 posts

Re: Qwen 3.8

#141

Qwen has set an excellent track record for architecting and releasing open-weight models that consumer-grade devices can run. What is needed the most right now is something similar to Bonsai 27B, with a modest memory footprint, but faster and more capable. On-device models can make up for intelligence by being faster, thinking longer, or doing more quick iteration rounds.

> What is needed the most right now is something similar to Bonsai 27B, with a modest memory footpint, but faster and more capable Yeah, that'd be neat, but that's not what this announcement is about at all: > With a massive 2.4T parameters

dont we all deem the ability to improve large models as the defacto capability to produce small ones?

Re: Qwen 3.8

#142
post #109

So are locally-runnable models frozen at Qwen 3.6 now :/

Is qwen 3.6 27b the best model you can run locally at the moment? Not that I have the VRAM for it, but just curious.

People have been able to run DeepSeek v4 flash with a high spec Mac.

Re: Qwen 3.8

#145

Earlier quoted context omitted.

That would go against everything that Dario believes in (note that I refer to the CEO and not the company; the staff at Anthropic are not so ridiculous). He believes in Anthropic being the sole arbiter of the forefront of this technology, because it is all too dangerous in the hands of anyone else.

I’ve seen no evidence that he believes in anything. He comes off as just another slimy would-be monopolist to me.

I see no evidence that any ceo retains anything but the desire to capitalize on their marketplace of ideas for their own benefit. Like wolves inn sheep clothing, they'll put on any skin suit to convince people to keep giving them money and power.

And it has nothing to do with the individual, from what I can tell, 70% of the population placed in their position would become the same type of uberpath.

Re: Qwen 3.8

#146

The "second only to Fable 5" comment is pretty telling here. I remember early on when a lot of naysayers were saying that Fable was barely an improvement on Opus. Like it or not, Anthropic have a genuine moat right now with that model, provided they continue to allow people to use it. It will be genuinely exciting when an open model is able to beat it.

> saying that Fable was barely an improvement on Opus. Like it or not, Anthropic have a genuine moat right now with that model,

What's more interesting is that Anthropic moat shrunk to just that model. There's zero reason to use any other model from Anthropic right now. And once they take Fable off subscription there will be zero reason to have Anthropic subscription.

Re: Qwen 3.8

#147

I predict that no one will use this and everyone will use Kimi K3.

I've been playing around with K3 a bunch, but the verbosity of the reasoning makes complete e2e agent work basically cost the same as other smaller models, and I'm not seeing a huge difference in quality, just a way longer e2e completion time.

Same problem with every chinese model currently, they overthink way too much and take too much tokens and time.

Re: Qwen 3.8

#148
Do this giant open-weight models have less active params and could be run on consumer hardware or no ?

Re: Qwen 3.8

#149

Qwen has set an excellent track record for architecting and releasing open-weight models that consumer-grade devices can run. What is needed the most right now is something similar to Bonsai 27B, with a modest memory footprint, but faster and more capable. On-device models can make up for intelligence by being faster, thinking longer, or doing more quick iteration rounds.

I can't really blame them that the biggest labs focused on trainig and realeasing huge models.

The niche for small models should be filled with medium sized labs doing distillations of the huge ones into consumer grade hardware runnable models and LORAs for the huge ones.

Re: Qwen 3.8

#150
The only problem I had with Qwen, fine tuning on Colab, it takes 31 t/s while Gemma 4 is around 9 t/s, otherwise, one of best local LLM
Post reply on HN