I will TLDR you on our thought process, research, training and benchmarks. 1. When running our quantized Qwen 3.8 27B instances we were very annoyed by random reasoning loops (in the paper bellow refered to as "overthinking errors". These random loops were persistent throughout medium and low reasoning settings. 2. We found a paper by Meta that's supposed to target this phenomenon in PTQ, but when used straight out o…
Re: Show HN: Swift-Qwen3.8-27B, -58.3% thinking, x1.95 speed, accuracy of xhigh
#11Nice. Going to dl and give it a whirl.