Strong vision and reasoning performance, and the 35-a3b model run s pretty ok on a 16gb GPU with some CPU layers.
Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
41–50 of 286 posts
Re: Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
#42Are there any non-Chinese open models that offer comparable performance?
What's your problem with Chinese LLMs?
Re: Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
#43Re: Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
#44What kind of hardware does HN recommend or like to run these models?
For fast inference, you’d be hard pressed to beat an Nvidia RTX 5090 GPU. Check out the HP Omen 45L Max: https://www.hp.com/us-en/shop/pdp/omen-max-45l-gaming-dt-gt2...
Re: Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
#45Earlier quoted context omitted.
Haiku 4.5 is a reasoning model. [0] [0]: https://www-cdn.anthropic.com/7aad69bf12627d42234e01ee7c3630... > Claude Haiku 4.5, a new hybrid reasoning large language model from Anthropic in our small, fast model class. > As with each model released by Anthropic beginning with Claude Sonnet 3.7, Claude Haiku 4.5 is a hybrid reasoning model. This means that by default the model will answer a query rapidly, but users have…
Sure, marketing people gonna market. But Haiku's 'extended thinking' mode is very different than the reasoning capabilities of Sonnet or Opus. I would absolutely believe mar-ticles that Qwen has achieved Haiku 4.5 'extended thinking' levels of coding prowess.
Oh HN never change.
Re: Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
#46https://unsloth.ai/docs/models/qwen3.5#qwen3.5-27b “ Qwen3.5-27B For this guide we will be utilizing Dynamic 4-bit which works great on a 18GB RAM”
Re: Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
#47Earlier quoted context omitted.
Sure, marketing people gonna market. But Haiku's 'extended thinking' mode is very different than the reasoning capabilities of Sonnet or Opus. I would absolutely believe mar-ticles that Qwen has achieved Haiku 4.5 'extended thinking' levels of coding prowess.
>Sure, marketing people gonna market. Oh HN never change.
Re: Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
#48...yeah I doubt it
Re: Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
#49I recently wrote a guide on getting: - llama.cpp - OpenCode - Qwen3-Coder-30B-A3B-Instruct in GGUF format (Q4_K_M quantization) working on a M1 MacBook Pro (e.g. using brew). It was bit finicky to get all of the pieces together so hopefully this can be used with these newer models. https://gist.github.com/alexpotato/5b76989c24593962898294038...
Re: Qwen3.5 122B and 35B models offer Sonnet 4.5 performance on local computers
#50I recently wrote a guide on getting: - llama.cpp - OpenCode - Qwen3-Coder-30B-A3B-Instruct in GGUF format (Q4_K_M quantization) working on a M1 MacBook Pro (e.g. using brew). It was bit finicky to get all of the pieces together so hopefully this can be used with these newer models. https://gist.github.com/alexpotato/5b76989c24593962898294038...