Earlier quoted context omitted.
Why would they use their most expensive model when sonnet or opus can do the job as well?
In my experience sonnet I have no reason to believe that the next generation won’t offer similar gains in verification, and there is some evidence to support that the cybersecurity implications are the result of exactly this expansion of ability.
Siccing Sonnet on a codebase or PR without guidance does indeed lead to worse results than using Opus, though.