Can we get the transcript of the session where you had Claude write this post? I want to understand the "why" behind this project, and whether a human being was at any point involved in the "blind, pre-registered evaluation of Grepathy against an honest baseline".
lol yes, agents ran the mechanics. I set the bars before any run and audited the keys. And the "why" behind the project is literally in the repo's .ai/why/ where you grep it if you would like.
I would encourage you to treat AI outputs with substantially more skepticism than you're giving them.