Live data from Hacker News

Qwen3.6-35B-A3B: Agentic coding power, now open to all

qwen.ai

281–290 of 563 posts

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#281
post #270

I've been running this on my laptop with the Unsloth 20.9GB GGUF in LM Studio: https://huggingface.co/unsloth/Qwen3.6-35B-A3B-GGUF/blob/mai... It drew a better pelican riding a bicycle than Opus 4.7 did! https://simonwillison.net/2026/Apr/16/qwen-beats-opus/

Oh that is pretty good! And the SVG one!

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#282

Earlier quoted context omitted.

How many people/hackernews can run a 397b param model at home? Probably like 20-30.

I can (barely, but sustainably) run Q3.5 397B on my Mac Studio with 256GB unified. It cost $10,000 but that's well within reach for most people who are here, I expect.

Hacker News moment

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#283

Earlier quoted context omitted.

This is just one model in the Qwen 3.6 series. They will most likely release the other small sizes (not much sense in keeping them proprietary) and perhaps their 122A10B size also, but the flagship 397A17B size seems to have been excluded.

How many people/hackernews can run a 397b param model at home? Probably like 20-30.

You can rent a cloud H200 with 140GB VRAM in a server with 256GB system ram for $3-4/hr.

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#284
post #73

I'm broadly curious how people are using these local models. Literally, how are they attaching harnesses to this and finding more value than just renting tokens from Anthropic of OpenAI?

It’s easy to find a combination of llama.cpp and a coding tool like OpenCode for these. Asking an LLM for help setting it up can work well if you don’t want to find a guide yourself. > and finding more value than just renting tokens from Anthropic of OpenAI? Buying hardware to run these models is not cost effective. I do it for fun for small tasks but I have no illusions that I’m getting anything superior to hosted m…

Oh lord, are the LLMs already replacing LLMs?

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#285
post #270

I've been running this on my laptop with the Unsloth 20.9GB GGUF in LM Studio: https://huggingface.co/unsloth/Qwen3.6-35B-A3B-GGUF/blob/mai... It drew a better pelican riding a bicycle than Opus 4.7 did! https://simonwillison.net/2026/Apr/16/qwen-beats-opus/

It's fascinating that a $999 Mac Mini (M4 32GB) with almost similar wattage as a human brain gets us this far.

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#286

Looking to move off ollama on Open Suse tumbleweed. Should I use brew to install llma.ccp or the zypper to install the tumbleweed package?

Why are you looking to move off Ollama? Just curious because I'm using Ollama and the cloud models (Kimi 2.5 and Minimax 2.7) which I'm having lots of good success with.

Ollama co mingles online and local models which defeats the purpose for me

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#287
My last laptop was a used 2012 T530.

My current is a used M1 MBP Pro with 16GB of ram.

I thought this was all I was ever going to need, but wanting to run really nice models locally has me thinking about upgrading.

Although, part of me wants to see how far I could get with my trusty laptop.

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#288
post #206

Earlier quoted context omitted.

We re-uploaded Gemma4 4 times - 3 times were due to 20 llama.cpp bug fixes, which we helped solve some as well. The 4th is an official Gemma chat template improvement from Google themselves, so these are out of our hands. All providers had to re-fix their uploads, so not just us. For MiniMax 2.7 - there were NaNs, but it wasn't just ours - all quant providers had it - we identified 38% of bartowski's had NaNs. Ours w…

Please publish sha256sums of the merged GGUFs in the model descriptions. Otherwise it's hard to tell if the version we have is the latest.

Why do you merge the GGUFs? The 50 GB files are more manageable (IMO) and you can verify checksums as you say.

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#289

My last laptop was a used 2012 T530. My current is a used M1 MBP Pro with 16GB of ram. I thought this was all I was ever going to need, but wanting to run really nice models locally has me thinking about upgrading. Although, part of me wants to see how far I could get with my trusty laptop.

Your current laptop is still a fine thin client. Unless you program in the woods, it's probably cheapest to build a home inference box and route it over Tailscale or something.

Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all

#290
post #270

I've been running this on my laptop with the Unsloth 20.9GB GGUF in LM Studio: https://huggingface.co/unsloth/Qwen3.6-35B-A3B-GGUF/blob/mai... It drew a better pelican riding a bicycle than Opus 4.7 did! https://simonwillison.net/2026/Apr/16/qwen-beats-opus/

But that you also gave a win to Qwen on flamingo is pretty outrageous! :)

Tthe right one looks much better, plus adding sunglasses without prompting is not that great. Hopefully it won't add some backdoor to the generated code without asking. ;)

Post reply on HN