Unsloth Dynamic 3.0 GGUFs
unsloth.ai
Unsloth Dynamic 3.0 GGUFs
1–10 of 125 posts
Re: Unsloth Dynamic 3.0 GGUFs
#2Re: Unsloth Dynamic 3.0 GGUFs
#3Re: Unsloth Dynamic 3.0 GGUFs
#4Hey Unsloth, your gguf are the first ones I look for when I want to download a gguf model. Today I was trying in fact to see, what's the smallest Qwen3.8-27B that I could run and get good results, say restricting it to 16GB of ram.. so I went, pick up the Qwen3.8-27B-UD-IQ2_XXS.gguf and them BAM, error on MTP... now I understand why after reading your announcement. Beyond the space saving, why removing the MTP? impro…
> We also removed the MTP module from smaller quants under UD-Q2_K_XL (8.37GB and lower) to converse around 500MB of disk space - you can use the Q4_0 MTP separate module if needed
Re: Unsloth Dynamic 3.0 GGUFs
#5 "We also made some smaller UD-1bit quants with UD-IQ1_S being 6.2GB (without MTP) which retain around 72% top-1% accuracy yet being 89% smaller"
This is crazy! But has anyone tried these lower quants on real projects?Re: Unsloth Dynamic 3.0 GGUFs
#6Re: Unsloth Dynamic 3.0 GGUFs
#7"We also made some smaller UD-1bit quants with UD-IQ1_S being 6.2GB (without MTP) which retain around 72% top-1% accuracy yet being 89% smaller" This is crazy! But has anyone tried these lower quants on real projects?
Re: Unsloth Dynamic 3.0 GGUFs
#8Re: Unsloth Dynamic 3.0 GGUFs
#9The new IQ4XS has been working pretty well so far on 4090 16gb.