Impressive work, love seeing tools that boost local LLM reliability without touching the model itself
I'm in the same boat, tuning models wasn't super interesting, though I might do a focused spike on behavior -focused fine tuning. But the harness matters almost more than the model in many cases.