Maybe the models or Cursor should warn you that you've got this vulnerability each time you use it.
Claude has learned how to jailbreak Cursor
21–30 of 40 posts
Re: Claude has learned how to jailbreak Cursor
#22Nothing to see here tbh. It's a very silly title for "claude sometimes writes shell scripts to execute commands it has been instructed aren't otherwise accessible"
We’ve reached a point where tools get hyped because they fail to follow instructions.
Re: Claude has learned how to jailbreak Cursor
#23Re: Claude has learned how to jailbreak Cursor
#24> Claude realized that I had to approve the use of such commands, so to get around this, it chose to put them in a shell script and execute the shell script. This sounds exactly like what anybody working sysops at big banks does to get around change controls. Once you get one RCE into prod, you’re the most efficient man on the block.
Re: Claude has learned how to jailbreak Cursor
#25Earlier quoted context omitted.
Gotta love the alarmist culture that surrounds these circles.
The same hype as the PlayStation being too powerful and potentially could be used by random countries to make nuclear weapons with a cluster of those.
Re: Claude has learned how to jailbreak Cursor
#26> we need to control the capabilities of software X > let's use blacklists, an idea conclusively proven never to work > blacklists don't work > Post title: rogue AI has jailbroken cursor
Re: Claude has learned how to jailbreak Cursor
#27What does "learned" mean in this context? LLMs don't modify themselves after training, do they?
Re: Claude has learned how to jailbreak Cursor
#28What a silly title, for a moment I thought Claude learned to exceed the Cursor quota limit... :s