Qwen 3.8 27B
huggingface.co
Qwen 3.8 27B
1–10 of 848 posts
Re: Qwen 3.8 27B
#2Really excited to see what people do with this. 3.7 27B was probably the best compromise between size and intelligence to run on consumer hardware
Re: Qwen 3.8 27B
#3This is one of the most important model releases since most use cases don't need SOTA/Frontier
If you want Qwen3.8-27B Serving Configs for the DGX Spark vLLM NVFP4 and RTX 4090 llama.cpp GGUF I added the setups here https://x.com/ErdalToprak/status/2088299678085308761?s=20
Re: Qwen 3.8 27B
#4These are massive improvements - and something you can actually run on a laptop.
Re: Qwen 3.8 27B
#5[deleted]
Re: Qwen 3.8 27B
#6My Strix Halo is about to go overdrive!
Re: Qwen 3.8 27B
#727b dense model at Opus 4.6 level
Opus at home
I hope there also will be a new ~10b variant
Re: Qwen 3.8 27B
#8Beats Opus 4.7 Max (w/ Claude Code) on DeepSWE (42.2 vs 40). Looks like Qwen's 27B models continue to pack some punch.
Unsloth's GGUF quants are up: https://huggingface.co/unsloth/Qwen3.8-27B-GGUF
Re: Qwen 3.8 27B
#9q4km is about 48 tps on a 4090. my llama.cpp params are --flash-attn on --parallel 1 --load-mode mmap
Re: Qwen 3.8 27B
#10I hope really badly that we'll get a new 35B A3B or similar MoE model!
I also miss the Qwen 3 Coder Next, which was 80B A3B, there are quite a few use cases where a non-dense model Also alternate link for viewing the images without signing in: https://xcancel.com/Alibaba_Qwen/status/2088280182356611304