Earlier quoted context omitted.
I generally agree that the harness isn't good, but it works and gets the job done and that seems to be the singular goal of the top 4 or 5 companies building them. We saw what Claude Code looks like inside, and it's objectively bad-to-mediocre work, but the takeaway seemed to be 'yeah but it works and they've got crazy revenue'. That's where we're at. The harness is kind of buggy. The LLM still wanders and cycles in…
> We saw what Claude Code looks like inside, and it's objectively bad-to-mediocre work Based on this, are there any open source harnesses that have objectively good-to-excellent work in their code?
It just has a sane workflow it's easy to use, doesn't bother you with 1000 questions if you allow this or that to run and generally it feels like the model is dumber and makes more mistakes since yesterday since I have to use claude code.