Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
1–10 of 268 posts
Re: Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
#2It's just a game, but I found the stats still interesting that I wanted to share back. Even with the warning up front, 1 in 3 threats were missed, and the history log above npm run commands seems to be typically ignored.
I also incorporated the feedback and insights from the previous HN thread, dns_snek's point about npm run in particular. Appreciate everyone who played and shared feedback!
Re: Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
#3It's been tried so many times before, and it never worked.
Re: Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
#4Re: Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
#5It's kinda funny there is still software coming out whose security model is "constantly ask the user for permission, and hope they never make a mistake". It's been tried so many times before, and it never worked.
Re: Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
#6This is a good case for custom harness/sandbox engineering.
Re: Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
#7Re: Humans missed 1 in 3 threats approving AI agent commands across 40k game runs
#8It’s simply a CYA click-thru by the model vendors so their lawyers can say “well you approved it this is on you” when AI does something stupid.