I've witnessed first-hand Claude (Opus 5 I think) doing some low-key "hacking" of sorts, when in a discovery phase of a task. It needed documentation for an API, but the documentation was behind a secure portal that it didn't have credentials to. However, it found an alternative route to the docs through an ensecured developer API. Not very advanced, but also not very far fetched to imagine it doing something a bit m…
Re: Anthropic says Claude AI hacked three organisations during cyber tests
#11[flagged]