Live data from Hacker News

Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

techcrunch.com

11–20 of 30 posts

Re: Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

#11
post #8

I tried browsing the Darpa challenge's website to know more, but I couldn't find any information. Could someone please post a link to a detailed description of the challenge?

It is basically computers playing Capture the Flag (CTF) against each other. They are given binary programs with security flaws. They need to identify the flaws automatically and develop a patch for their own system. At the same they go out to crash the other teams. Normally humans do this, but the darpa challenge was to have computer systems do it autonomously.

Re: Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

#12
post #8

I tried browsing the Darpa challenge's website to know more, but I couldn't find any information. Could someone please post a link to a detailed description of the challenge?

It is basically computers playing Capture the Flag (CTF) against each other. They are given binary programs with security flaws. They need to identify the flaws automatically and develop a patch for their own system. At the same they go out to crash the other teams. Normally humans do this, but the darpa challenge was to have computer systems do it autonomously.

Re: Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

#13

The Mayhem is also competing in the CTF.

It's not doing that hot - currently last place, but not very far back in terms of points.

However (and impressively), it did patch at least one bug in a task (LEGIT_00007) before any other human team did.

Re: Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

#14
Hey guys, member of the (currently unverified) third place team, Shellphish. If anyone has any questions, I (or another member of my team) would be glad to answer them. We'll also be giving a talk at DEF CON on Sunday after the CTF ends, where we'll be open sourcing our CRS!

Re: Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

#15
post #8

I tried browsing the Darpa challenge's website to know more, but I couldn't find any information. Could someone please post a link to a detailed description of the challenge?

https://www.cybergrandchallenge.com/tech

Includes a link to the github for the challenge framework.

Re: Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

#16

Hey guys, member of the (currently unverified) third place team, Shellphish. If anyone has any questions, I (or another member of my team) would be glad to answer them. We'll also be giving a talk at DEF CON on Sunday after the CTF ends, where we'll be open sourcing our CRS!

What kind of AI was involved in your and competitors systems?

Re: Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

#17

Hey guys, member of the (currently unverified) third place team, Shellphish. If anyone has any questions, I (or another member of my team) would be glad to answer them. We'll also be giving a talk at DEF CON on Sunday after the CTF ends, where we'll be open sourcing our CRS!

Whats your view on complete automation vs human assisted automation? Which one is better to focus building on for a 5 year timeline?

Re: Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

#19
I thought that the production of the competition was extraordinary. Seeing everything lit up on stage was straight out of a movie (in a good way). I thought that the event itself at Defcon was super weird, though. A lot of people, myself included, assumed that the event was going to be more real-time. In reality, the servers had been competing for hours already.

That being said, huge props to these amazing teams. It was so fascinating to see how each system reacted to the same situations and then either hunkered down to protect itself or go on the offensive. Really amazing stuff.

Re: Carnegie Mellon’s Mayhem AI Wins DARPA’s Cyber Grand Challenge

#20

Hey guys, member of the (currently unverified) third place team, Shellphish. If anyone has any questions, I (or another member of my team) would be glad to answer them. We'll also be giving a talk at DEF CON on Sunday after the CTF ends, where we'll be open sourcing our CRS!

Can you explain how this particular CTF work and how the system in general work against adversary? The article said insecure code and code filled with bugs are constantly being fed to the system. I don't really get it.
Post reply on HN