Live data from Hacker News

Qwen3.8-Max: A New Bar for Coding and Cowork

qwen.ai

591–600 of 652 posts

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#591
post #554

Earlier quoted context omitted.

Not sure what you are trying to say but thanks for reminding me about qm, it has some cool features. My MindRoot framework had some of qm's main features awhile ago though. I do try to use it to build solutions by customizing rather than from scratch when possible. But a lot of clients or potential clients don't like that idea even though it's MIT. They want some that has clearly been invented by them as a new progra…

Why are you looking for projects on upwork if you know they don’t pay well and that you don’t have leverage? That sounds like a pretty bad place where to find contracting projects

Please recommend somewhere else. I have started on LinkedIn but especially without much time and having so much experience being shortish Upwork contracts it's a hard sell without lying on my resume and so I am having to be careful where I spend my time.

The biggest advantage of Upwork is I have a track record there and there are tons of RFPs. Other places usually don't have RFPs and are going to look down on the Upwork record.

But I appreciate any ideas for places or people to look for work. My email is in my profile.

I am not sure about funding for rent if I don't pick something up this week.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#592

Earlier quoted context omitted.

Qwen-3.6-35B-A3B was our "gateway drug" into switching our organisation to agent/harness-first coding. Particularly, I had one team member who was extremely sceptical of AIs/LLMs/harnesses and refused to use them. One day he said "Well, I have an RTX 5090 doing nothing... should I try to get something up on it?" and a few minutes later he had 3.6-35B loaded up, running OpenCode. It continues to be a workhorse to this…

I’ve run 3.6-27B and 3.6-35B on 32GB locally for a lot of bulk non-code tasks. Let it run overnight and wake up to millions of output tokens worth of results without data having left my house, all for the price of electricity. I haven’t found it very useful for code. It can do some code, but I’ve tried a dozen different quants and context lengths and the output is always bad enough that it has to be discarded for any…

automatic coding is a flow thing with 3.6, but I agree it isnt fully there, it does need specific per function instructions to maintain coherence, with a few exceptions Im unsure hold in general which I got via an experimental harness.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#593

Earlier quoted context omitted.

Yeah, a good rule of thumb is that the weights take up ~100% of the size of the model, so 100B bytes (8-bit quant) would be, well, 100GB and a 4-bit quant would be half that.

I don't know if that's, well, a rule of thumb, it might be, well, straight multiplication.

Well, yeah, the straight multiplication, er, well, "follows" from the rule of thumb. Hope that helps!

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#594

> How Qwen Cloud handles your data during inference? > Qwen Cloud does not use your API inputs or outputs to train or improve models. > Learn more about how your data is handled during the inference process from Zero data retention. If this is true, this is a big deal for me, but unfortunately I cannot find anything in their legal agreements, so this marketing sentence is worthless.

It must be considered false for all tech providers unless there is a verification mechanism. There will be big money for whoever sorts that out

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#595

Earlier quoted context omitted.

I don't know if that's, well, a rule of thumb, it might be, well, straight multiplication.

Well, yeah, the straight multiplication, er, well, "follows" from the rule of thumb. Hope that helps!

Saying, well a rule of thumb is, well, 100 billion bytes is a 100 gigabytes, is well, not a rule of thumb. It is, well, just the common definition.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#596

Earlier quoted context omitted.

Qwen-3.6-35B-A3B was our "gateway drug" into switching our organisation to agent/harness-first coding. Particularly, I had one team member who was extremely sceptical of AIs/LLMs/harnesses and refused to use them. One day he said "Well, I have an RTX 5090 doing nothing... should I try to get something up on it?" and a few minutes later he had 3.6-35B loaded up, running OpenCode. It continues to be a workhorse to this…

> Qwen-3.6-35B-A3B The A3B models are super fast but I found the A3B Q4 model ran in circles a lot and ended up taking longer to complete tasks that 27B Q6 because it kept having to redo/rethink/fix something. I was writing extensive prompts to rein it in and it would still ignore basic directives like "never force push on the repo, ask me instead". I ended up switching back to 27B after about a week of frustration a…

[deleted]

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#597

Earlier quoted context omitted.

Qwen-3.6-35B-A3B was our "gateway drug" into switching our organisation to agent/harness-first coding. Particularly, I had one team member who was extremely sceptical of AIs/LLMs/harnesses and refused to use them. One day he said "Well, I have an RTX 5090 doing nothing... should I try to get something up on it?" and a few minutes later he had 3.6-35B loaded up, running OpenCode. It continues to be a workhorse to this…

> Qwen-3.6-35B-A3B The A3B models are super fast but I found the A3B Q4 model ran in circles a lot and ended up taking longer to complete tasks that 27B Q6 because it kept having to redo/rethink/fix something. I was writing extensive prompts to rein it in and it would still ignore basic directives like "never force push on the repo, ask me instead". I ended up switching back to 27B after about a week of frustration a…

I'm playing/experimenting with a harness and just tested how well various models follow the instructions, and how they react to the tool claiming a local temperature of 72°C

here is how qwen3.6-27b reacted:

https://pastebin.com/srf7gjfy

try to count the number of times it "thinks" okay ready, just say the thing, no wait but what if...

this isn't (a mimicry of) thinking, this is (a mimicry of) insecurity/fear

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#598

Earlier quoted context omitted.

I'm running Qwen 27B no problem with an AMD 9070XT + 24gb DDR5 ram. Does basic web search for me (tool call with tavily, costs nothing I get 1000 searches a month) and is great for creative writing (primarily breaking writer's block). Until the recent surge in ram costs, that wouldn't be hard to do. I built the computer for ~$1600 a year ago.

I am getting ~13-15 tps with my 9070XT for the 27B (~35tps for the 35B-A3B), but I think for me the main bottleneck is the 64gb of DDR4 3600 memory. What kinda speeds are you getting with what speed of DDR5?

I’m a little more novice than a lot of the people on this site so take my response with a grain of salt.

The wall I keep hitting is I can run models like I described (Q3-4 usually), but it’s very sensitive to context. Once I start getting past 7.5k or so it can really fall apart. Sometimes before that. It just really depends.

If I run smaller ones that offload less to ram, they stay somewhat coherent but don’t quite do what I want them to do.

Your token speeds are not that much slower than mine. I imagine part of it is I’m not fine-tuning it very well. On a good day I’ll get like…15tps.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#599
post #301

Earlier quoted context omitted.

With laptop being ...?

It's just an M4 MacBook Air with 16gb ram – probably incapable of running models itself. I actually run the models on my Mac Studio which is an M1 Ultra with 64gb, and oh-my-pi on my laptop is configured to use the models over the local network.

Yeah I have 24 GB Ram on my M4 and am disappointed in the capabilities. You really need a ton of RAM. I can run some basic 8B models fine but they dont do well at all in basic coding stuff I've thrown at them.

Re: Qwen3.8-Max: A New Bar for Coding and Cowork

#600

They've also announced Qwen3.8-27B being released open-weight next week. Qwen3.6-27B is widely regarded as one of the best local models, especially since nothing else comes close to it, that isn't benchmaxxed, without being significantly larger. If 3.8 truly improves upon it that would be awesome.

Qwen3.8-Max is the first in Qwen-Max series to be open-weight as well. Kimi K3, GLM 5.2 and now Qwen3.8-Max - open weight models. DeepSeek V4 Flash outperforming Gemini 3.1 pro, probably DeepSeek V4 Pro update is also coming soon Chinese labs are cooking very hard. US closed weight labs are probably hard time to resist not calling Washington DC for more AI regulations

I dont think they have a hard time resisting they did it a while ago, my company mandated everyone delete anything Chinese or Chinese derived back in April I think for no reason than "unsafe"
Post reply on HN