nb: current models (e.g. GPT 5.6 Sol) are very good at long horizon tasks they no longer need crutches or rube goldberg machines to keep them going minimal agent harness is just a loop that loops until no more tool calls are coming GPT 5.6 Sol continues to drive the loop until the task is done or it decides that it wants to present the user with information at that point it is probably good to not automatically conti…
In my daily work, I found expeciall terra and sol now stopping every few rounds again, telling me the tak is done. I even had them create a detailled plan - and told them to finish "end to end" - and they appruptly stop after the plan. Because they interpret this as finished. Even if the DOD is clearly not "finish the plan". The new models are shite (pardon my French), when it comes to long running tasks and I find m…
but I don't doubt that you're seeing this behaviour, ty for sharing!