Live data from Hacker News

Which small model is best for fine-tuning? We tested 12 of them on 8 tasks

distillabs.ai

1–2 of 2 posts

Re: Which small model is best for fine-tuning? We tested 12 of them on 8 tasks

#2
We benchmarked which small language models are most tunable and which deliver best performance after fine-tuning. Tested 12 models (Qwen, Llama, Gemma, Granite, SmolLM) on 8 tasks.

TL;DR Qwen3 family is the best overall choice, small Llamas improve the most after fine-tuning.