Bring it on! Hoping that they release smaller sizes of Qwen3.8. I use the 35B MoE and 27B dense models locally and most of the time I don’t need to reach out to Claude. Extremely useful specially when requests include sensitive and/or personal data
Qwen 3.8
61–70 of 793 posts
Re: Qwen 3.8
#62Earlier quoted context omitted.
In China, you can’t officially use US APIs. The world saw a taste of this with Fable, but in China, this has been the situation all along. So it’s not a surprise why open weights are so cherished. As frontier models continue to block everyday individuals from securing their own codebase, I expect the adoption and usage of open weights to continue. As an example, HuggingFace recently was investigating a security incid…
How does this explain open weights? They could easily take the same closed route like their American friends
So it seems like it's very important to them that people use the Qwen app and they're willing to pay a lot of money for that. Presumably someone thought that keeping their best models closed would drive more business to them (as the sole provider) but then they discovered that closed releases mostly get ignored unless they're really good. (See also: People who think that Chinese AI companies are required to release weights as a matter of policy, because the closed ones hardly ever show up in the news.) Releasing weights for Qwen 3.8 at least lets them get some of that "pretty good for the price of free" media buzz.
Re: Qwen 3.8
#63Re: Qwen 3.8
#64Earlier quoted context omitted.
It's hard to say what their motivation is. The Chinese firms seem to be working hard to commoditize intelligence which may be the most effective way to debase American frontier labs. And yeah: it also happens to be really good for humanity.
> It's hard to say what their motivation is. Not that hard to say IMO, they basically see models becoming a commodity and see value in the applications on top of them. So if Alibaba Cloud is the best place to build applications on top of Qwen, why not give the model itself away?
Yeah. Its a bit like the "open core" model in open source.
Re: Qwen 3.8
#65Earlier quoted context omitted.
> It's hard to say what their motivation is. Not that hard to say IMO, they basically see models becoming a commodity and see value in the applications on top of them. So if Alibaba Cloud is the best place to build applications on top of Qwen, why not give the model itself away?
> Not that hard to say IMO, Unless you work there, your opinions are guesses, and parent is saying we cannot know, which remains true even with your guesses :)
And my point is while we cannot know, it's not hard to make an informed guess as to their motivations i.e. there's some fairly obvious motivations here, not sure what yours is?
Re: Qwen 3.8
#66Re: Qwen 3.8
#67Earlier quoted context omitted.
I think everyone is hoping this! It would be great if they'd release an MoE model somewhere between the 35B size of 3.6 and the 122B version of 3.5 - it could be a great balance of speed and ability for people with reasonably powerful but not insane home computers.
> the 122B version of 3.5 Yeah, this is what I'm holding out for, the NVFP4 variant of 3.5 122B is blazing fast with reasonable quality and even with max context fits perfectly within 96GB.
Re: Qwen 3.8
#68Will probably be at Opus 4.8 level, and I find it pretty big deal because of Deepseek price...
Re: Qwen 3.8
#69I assume that this announcement has been prompted by that of Moonshot AI, which has just announced a 2.8T parameter open-weights LLM, Kimi K3, to be published on Huggingface by 27 July. Now the response of Alibaba is that they will also publish soon a big open weights LLM, the 2.4T parameter Qwen 3.8. I wonder if Alibaba has always planned to make this big LLM open weights, or they have chosen to do this now, to bett…
I would rather see them releasing 3.7-27B, 3.7-122B or their 3.8 versions. Qwen/QwQ were always about the best available local inference at home.
Re: Qwen 3.8
#70Earlier quoted context omitted.
> the 122B version of 3.5 Yeah, this is what I'm holding out for, the NVFP4 variant of 3.5 122B is blazing fast with reasonable quality and even with max context fits perfectly within 96GB.
Sweet. Are you running a Mac Studio Ultra or 4x3090?