They've also announced Qwen3.8-27B being released open-weight next week. Qwen3.6-27B is widely regarded as one of the best local models, especially since nothing else comes close to it, that isn't benchmaxxed, without being significantly larger. If 3.8 truly improves upon it that would be awesome.
Qwen3.8-Max: A New Bar for Coding and Cowork
41–50 of 652 posts
Re: Qwen3.8-Max: A New Bar for Coding and Cowork
#42the benchmark I trust most is whether the model can explain its own pricing page without getting confused
Not even humans can do that, you're literally asking for something beyond AGI
Re: Qwen3.8-Max: A New Bar for Coding and Cowork
#43Earlier quoted context omitted.
Qwen3.6-35B is my daily driver for AI, and what convinced me to cancel my Claude subscription back in April. The Qwen3.6 line is easily the best local model I've tried, and I've tried a lot. I've got it diligently grinding away on my laptop right now, reviewing and fixing some bugs in my F# code.
35B MoE is certainly a good and fast local model. I find 27B dense to be quite a bit smarter, so I daily drive that. I wish there was a ~100B MoE with maybe 10B active. It would be super smart and fast!
Re: Qwen3.8-Max: A New Bar for Coding and Cowork
#44Earlier quoted context omitted.
... other comments were right, this is the full qwen3.8-max model, two weeks ago was the qwen3.8-max-preview release. Here's a pelican I just got out of the new model. It took 11 minutes and forgot the wheels! https://tools.simonwillison.net/markdown-svg-renderer#url=ht... (scroll to bottom) The reasoning trace is pretty great: > More additions: basket with fish in it? Cute detail — a fish poking out of a basket on t…
Wow, bike geometry is really good! Except for the missing wheels lol
Re: Qwen3.8-Max: A New Bar for Coding and Cowork
#45Once OpenAI and Anthropic are public, every such announcement will become a reliable sell signal
Re: Qwen3.8-Max: A New Bar for Coding and Cowork
#46It seems this is the only mention of cost? > Qwen3.8-Max comes with the official support for reasoning_effort, which can be used to adjust reasoning depth and control cost: > xhigh (default): for complex tasks demanding thorough analysis > medium: balancing accuracy and speed > low: efficient reasoning optimizing for speed and cost I hope this is significantly cheaper. I've been loving Deepseek for it's nearly free u…
The linked qwencloud page has pricing; it's $2/6. https://www.qwencloud.com/models/qwen3.8-max
Re: Qwen3.8-Max: A New Bar for Coding and Cowork
#47They don't explain how successful that went but it's a bit hilarious seen that an Anthropic dev explained that it's been 15 days Claude was hard at work --with nothing to show yet-- trying to rewrite itself in another language.
"You rewrite Claude Code, we rewrite oh-my-pi."
"You're nowhere after 15 days, we do it in 10."
Sure, it's apples to oranges and all that. But part of me thinks they know fully well what they did there.
Re: Qwen3.8-Max: A New Bar for Coding and Cowork
#48Lmao I love their video with the idea that people will be able to do their hobbies while ai does their job. Surely Alibaba is leading by example here by reducing work hours per week while keeping pay the same right? Right?
That's the thing. Wny are companies like OpenAI/Anthropic/Alibaba/Kimi/Deepseek still hiring SWEs if their models have become so good?
Re: Qwen3.8-Max: A New Bar for Coding and Cowork
#49Earlier quoted context omitted.
Qwen3.6-35B is my daily driver for AI, and what convinced me to cancel my Claude subscription back in April. The Qwen3.6 line is easily the best local model I've tried, and I've tried a lot. I've got it diligently grinding away on my laptop right now, reviewing and fixing some bugs in my F# code.
compared to claude - how 'fast' is it in terms of throughput on your laptop?
Re: Qwen3.8-Max: A New Bar for Coding and Cowork
#50Earlier quoted context omitted.
Qwen3.6-35B is my daily driver for AI, and what convinced me to cancel my Claude subscription back in April. The Qwen3.6 line is easily the best local model I've tried, and I've tried a lot. I've got it diligently grinding away on my laptop right now, reviewing and fixing some bugs in my F# code.
What laptop?