I've been running this on my laptop with the Unsloth 20.9GB GGUF in LM Studio: https://huggingface.co/unsloth/Qwen3.6-35B-A3B-GGUF/blob/mai... It drew a better pelican riding a bicycle than Opus 4.7 did! https://simonwillison.net/2026/Apr/16/qwen-beats-opus/
Qwen3.6-35B-A3B: Agentic coding power, now open to all
281–290 of 563 posts
Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all
#282Earlier quoted context omitted.
How many people/hackernews can run a 397b param model at home? Probably like 20-30.
I can (barely, but sustainably) run Q3.5 397B on my Mac Studio with 256GB unified. It cost $10,000 but that's well within reach for most people who are here, I expect.
Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all
#283Earlier quoted context omitted.
This is just one model in the Qwen 3.6 series. They will most likely release the other small sizes (not much sense in keeping them proprietary) and perhaps their 122A10B size also, but the flagship 397A17B size seems to have been excluded.
How many people/hackernews can run a 397b param model at home? Probably like 20-30.
Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all
#284I'm broadly curious how people are using these local models. Literally, how are they attaching harnesses to this and finding more value than just renting tokens from Anthropic of OpenAI?
It’s easy to find a combination of llama.cpp and a coding tool like OpenCode for these. Asking an LLM for help setting it up can work well if you don’t want to find a guide yourself. > and finding more value than just renting tokens from Anthropic of OpenAI? Buying hardware to run these models is not cost effective. I do it for fun for small tasks but I have no illusions that I’m getting anything superior to hosted m…
Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all
#285I've been running this on my laptop with the Unsloth 20.9GB GGUF in LM Studio: https://huggingface.co/unsloth/Qwen3.6-35B-A3B-GGUF/blob/mai... It drew a better pelican riding a bicycle than Opus 4.7 did! https://simonwillison.net/2026/Apr/16/qwen-beats-opus/
Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all
#286Looking to move off ollama on Open Suse tumbleweed. Should I use brew to install llma.ccp or the zypper to install the tumbleweed package?
Why are you looking to move off Ollama? Just curious because I'm using Ollama and the cloud models (Kimi 2.5 and Minimax 2.7) which I'm having lots of good success with.
Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all
#287My current is a used M1 MBP Pro with 16GB of ram.
I thought this was all I was ever going to need, but wanting to run really nice models locally has me thinking about upgrading.
Although, part of me wants to see how far I could get with my trusty laptop.
Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all
#288Earlier quoted context omitted.
We re-uploaded Gemma4 4 times - 3 times were due to 20 llama.cpp bug fixes, which we helped solve some as well. The 4th is an official Gemma chat template improvement from Google themselves, so these are out of our hands. All providers had to re-fix their uploads, so not just us. For MiniMax 2.7 - there were NaNs, but it wasn't just ours - all quant providers had it - we identified 38% of bartowski's had NaNs. Ours w…
Please publish sha256sums of the merged GGUFs in the model descriptions. Otherwise it's hard to tell if the version we have is the latest.
Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all
#289My last laptop was a used 2012 T530. My current is a used M1 MBP Pro with 16GB of ram. I thought this was all I was ever going to need, but wanting to run really nice models locally has me thinking about upgrading. Although, part of me wants to see how far I could get with my trusty laptop.
Re: Qwen3.6-35B-A3B: Agentic coding power, now open to all
#290I've been running this on my laptop with the Unsloth 20.9GB GGUF in LM Studio: https://huggingface.co/unsloth/Qwen3.6-35B-A3B-GGUF/blob/mai... It drew a better pelican riding a bicycle than Opus 4.7 did! https://simonwillison.net/2026/Apr/16/qwen-beats-opus/
Tthe right one looks much better, plus adding sunglasses without prompting is not that great. Hopefully it won't add some backdoor to the generated code without asking. ;)