Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
modelscope.cn
Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
1–10 of 178 posts
Re: Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
#2Can you share the source for the parameter count (125B A6B)? I didn't see it anywhere in the page.
Re: Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
#3Can you share the source for the parameter count (125B A6B)? I didn't see it anywhere in the page.
It was in description under the countdown initially, but was quickly removed.
It also said 51B of n-grams and new attention (IIRC it said "Qwen Sparse Attention").
edit: here's a random screenshot https://x.com/AiBattle_/status/2092210011858460819/photo/1
Re: Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
#4Re: Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
#5Alibaba is giving sleepless nights to the tech giants
Re: Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
#6lol blocked with dns4eu
what a joke this resolver has become
Re: Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
#7Wow. I wasn't expecting this. I thought they were going to do a 35B model instead.
Re: Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
#8I hope there is going to be a free endpoint... Unlike 35B-A3B, I am nowhere close to running it locally
Re: Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
#9gg
Re: Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
#10> We are releasing these architectural improvements ahead of time so that the community can prepare for the upcoming full family of Qwen4 models.
That gives me hope that "full family" means it will include smaller models like 4B.