1. It is what you do after the prompt: scan diffs in real time and hit the escape key at the slightest sign of trouble. 2. Deep questioning. Constantly probing the assistant: what does it think it is trying to achieve, why did it just make decision X, is there a better way, what does it think the current constraint is? 3. Fighting drift. Knowing that the model will always try to regress to the mean of the training co…
[flagged]