Earlier quoted context omitted.
It should have the same flow as reviewing PRs from humans.
Who really truly enjoys that and doesn't see it as a chore? I find the real way to review other people's code is to program with it and then I start seeing where the problems are much more clearly. I would do a review and spot nothing important then start working on my own follow-on change and immediately run into issues.
I think it becomes a chore when there are too many trivial mistakes, and you feel like your time would have been better spent writing it yourself. As models and agent frameworks improve I see this happening less and less.