Earlier quoted context omitted.
In my experience with Qwen3.6 35B-A3B, disabling thinking made the model generate inaccurate replies. Ask it for the recipe of egg salad and it gives you the recipe of an omelette. Did I miss something, is it possible to have that model be reliable without thinking?
The official docs recommend different sampler options for thinking off. Have you tried adjusting those?
I used the settings recommended on huggingface/unsloth's model page + whatever suggestion from various LLMs - didn't research too much myself which setting did what.