Live data from Hacker News

Qwen3.8-Max: A New Bar for Coding and Cowork

qwen.ai

41–50 of 652 posts

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#41

They've also announced Qwen3.8-27B being released open-weight next week. Qwen3.6-27B is widely regarded as one of the best local models, especially since nothing else comes close to it, that isn't benchmaxxed, without being significantly larger. If 3.8 truly improves upon it that would be awesome.

If they trained it well, and can do computer use, it will be a new era. Companies can keep PCs, put Qwen 3.8 27b on it and get rid of the employees, lol...

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#42
post #30

the benchmark I trust most is whether the model can explain its own pricing page without getting confused

Not even humans can do that, you're literally asking for something beyond AGI

Bistromathics

https://www.hhgproject.org/entries/bistromathics.html

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#43

Earlier quoted context omitted.

Qwen3.6-35B is my daily driver for AI, and what convinced me to cancel my Claude subscription back in April. The Qwen3.6 line is easily the best local model I've tried, and I've tried a lot. I've got it diligently grinding away on my laptop right now, reviewing and fixing some bugs in my F# code.

35B MoE is certainly a good and fast local model. I find 27B dense to be quite a bit smarter, so I daily drive that. I wish there was a ~100B MoE with maybe 10B active. It would be super smart and fast!

There was a 3.5 122B 10A release -

https://huggingface.co/Qwen/Qwen3.5-122B-A10B

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#44
post #29

Earlier quoted context omitted.

... other comments were right, this is the full qwen3.8-max model, two weeks ago was the qwen3.8-max-preview release. Here's a pelican I just got out of the new model. It took 11 minutes and forgot the wheels! https://tools.simonwillison.net/markdown-svg-renderer#url=ht... (scroll to bottom) The reasoning trace is pretty great: > More additions: basket with fish in it? Cute detail — a fish poking out of a basket on t…

Wow, bike geometry is really good! Except for the missing wheels lol

Do pelican bikes need wheels? They've got wings, after all... I think Qwen is on to something here.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#45
post #20

Once OpenAI and Anthropic are public, every such announcement will become a reliable sell signal

It's not so simple, if such a headline can get them closer to the regulatory capture they want to lock in American businesses and forbid them from using Chinese AI.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#46
post #9

It seems this is the only mention of cost? > Qwen3.8-Max comes with the official support for reasoning_effort, which can be used to adjust reasoning depth and control cost: > xhigh (default): for complex tasks demanding thorough analysis > medium: balancing accuracy and speed > low: efficient reasoning optimizing for speed and cost I hope this is significantly cheaper. I've been loving Deepseek for it's nearly free u…

The linked qwencloud page has pricing; it's $2/6. https://www.qwencloud.com/models/qwen3.8-max

Thanks, totally missed that!

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#47
> In this case, Qwen3.8-Max was asked to create the oh-my-cli project from scratch and, over a 10+ day long-horizon autonomous coding run, build a self-evolving harness.

They don't explain how successful that went but it's a bit hilarious seen that an Anthropic dev explained that it's been 15 days Claude was hard at work --with nothing to show yet-- trying to rewrite itself in another language.

"You rewrite Claude Code, we rewrite oh-my-pi."

"You're nowhere after 15 days, we do it in 10."

Sure, it's apples to oranges and all that. But part of me thinks they know fully well what they did there.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#48
post #7
post #4

Lmao I love their video with the idea that people will be able to do their hobbies while ai does their job. Surely Alibaba is leading by example here by reducing work hours per week while keeping pay the same right? Right?

That's the thing. Wny are companies like OpenAI/Anthropic/Alibaba/Kimi/Deepseek still hiring SWEs if their models have become so good?

https://en.wikipedia.org/wiki/Jevons_paradox

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#49
post #34

Earlier quoted context omitted.

Qwen3.6-35B is my daily driver for AI, and what convinced me to cancel my Claude subscription back in April. The Qwen3.6 line is easily the best local model I've tried, and I've tried a lot. I've got it diligently grinding away on my laptop right now, reviewing and fixing some bugs in my F# code.

compared to claude - how 'fast' is it in terms of throughput on your laptop?

I use it with a strix halo server. 35B runs stupidly fast. 27B is about 700 TPS prefill and 30 TPS token generation. Which interestedly is about what Kimi K3 gives me depending on provider.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#50

Earlier quoted context omitted.

Qwen3.6-35B is my daily driver for AI, and what convinced me to cancel my Claude subscription back in April. The Qwen3.6 line is easily the best local model I've tried, and I've tried a lot. I've got it diligently grinding away on my laptop right now, reviewing and fixing some bugs in my F# code.

What laptop?

It's just a MacBook Air with an M4, cheap and nothing special. I host Qwen on my Mac Studio, an M1 with 64gb ram. The model uses around 20-25gb ram depending on what it's doing.
Post reply on HN