Live data from Hacker News

Qwen3.8-Max: A New Bar for Coding and Cowork

qwen.ai

221–230 of 652 posts

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#221

They've also announced Qwen3.8-27B being released open-weight next week. Qwen3.6-27B is widely regarded as one of the best local models, especially since nothing else comes close to it, that isn't benchmaxxed, without being significantly larger. If 3.8 truly improves upon it that would be awesome.

Having invested in a machine with 128GB of RAM, I would love seeing something a bit larger than 27B / 35B, possibly a 54B dense model or 70B MoE would be much closer to the Qwen 3.8 Max experience.

All of us with a 96gb rtx 6000 would love to see a 70b moe. Maybe they are waiting for OpenAI and Anthropic to IPO so they can short their stock and release. Local LLM is going to get very interesting in the next 2 years.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#222
post #216

Earlier quoted context omitted.

Qwen-3.6-35B-A3B was our "gateway drug" into switching our organisation to agent/harness-first coding. Particularly, I had one team member who was extremely sceptical of AIs/LLMs/harnesses and refused to use them. One day he said "Well, I have an RTX 5090 doing nothing... should I try to get something up on it?" and a few minutes later he had 3.6-35B loaded up, running OpenCode. It continues to be a workhorse to this…

[flagged]

A local model needs 0 investment and 0 commitment, takes literal minutes to get started (especially if you have someone who is into that stuff showing you the ropes) and if you end up disliking the experience of using AI you can just `rm -fr` it and forget the whole thing existed.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#223
post #147

It was a matter of time for China to catch up with the US. In terms of infrastructure, manufacturing, and engineering workforce, China has the upperhand and I foresee them becoming the SOTA leaders. Maybe if the US wasn't so busy gatekeeping and keeping things proprietary, they would've had more trust from the open source community.

First and foremost, China isn't energy-limited like the US is. A modern grid, lost of cheap power coming online every day. The worlds most advance ultra-high voltage transport links are tying the deserts in the west to the cities in the south. It probably sucks for you if the proposed powerline goes through your house, but the need of the many come first. Having cheap and plentiful power is a huge market advantage. I…

> but the need of the many come first.

The need of the rich, you mean. In these kinds of situations it's never about the many, if it was about them different choices would be made.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#224
post #205

Earlier quoted context omitted.

Some advice I got from another HN Mac user was to run local models in energy saver mode. You'll get slightly reduced tokens, but the laptop won't overheat and the fans won't go wild.

Oh. I've been using an icepack under my laptop to keep mine cool. I'm watching it with llamatop to see if the GPU is actually active or not, aw activity monitor wasn't showing me what I wanted.

That's exactly how i fried a laptop. The condensation killed it.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#225
post #216

Earlier quoted context omitted.

Qwen-3.6-35B-A3B was our "gateway drug" into switching our organisation to agent/harness-first coding. Particularly, I had one team member who was extremely sceptical of AIs/LLMs/harnesses and refused to use them. One day he said "Well, I have an RTX 5090 doing nothing... should I try to get something up on it?" and a few minutes later he had 3.6-35B loaded up, running OpenCode. It continues to be a workhorse to this…

[flagged]

Maybe you're anti-AI because you're really anti-outsourcing your thinking to some remote corporation you don't control?

That's one of my main issues with AI anyways, the thought of having all my data go through some sketchy foreign (to me) entity with questionable motives and under a questionable regime.

Local AI solves for all of those.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#226
post #59

Earlier quoted context omitted.

what hardware do you use or recommend for this? never heard of it until today.

Strix Halo is the unified memory platform from AMD. Similar to the DGX Spark from NVIDIA or the M series Macs. I personally have the Framework Desktop, but there's also systems from other brands like Bosgame

You can also get it in a laptop form factor that feels like a MBP with a nicer keyboard if you get an HP Zbook G1A!

Huge fan of that thing, it's th e Linux MBP I've always wanted.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#227

They've also announced Qwen3.8-27B being released open-weight next week. Qwen3.6-27B is widely regarded as one of the best local models, especially since nothing else comes close to it, that isn't benchmaxxed, without being significantly larger. If 3.8 truly improves upon it that would be awesome.

Qwen3.8-Max is the first in Qwen-Max series to be open-weight as well.

Kimi K3, GLM 5.2 and now Qwen3.8-Max - open weight models.

DeepSeek V4 Flash outperforming Gemini 3.1 pro, probably DeepSeek V4 Pro update is also coming soon

Chinese labs are cooking very hard. US closed weight labs are probably hard time to resist not calling Washington DC for more AI regulations

Post reply on HN