Live data from Hacker News

OpenAI Preparedness Challenge

openai.com

41–50 of 160 posts

Re: OpenAI Preparedness Challenge

#42

This kinda of crowdsourcing just feels.... f'ing weird man. It's like if, after the 1993 WTC bombing, but before 9/11, the FBI and NY Port Authority went around asking people how they would attack NYC if they were to become terrorists and... then how they would suggest detecting and stopping said attack. And please be as detailed as possible. Leave your name and phone number. Best answer gets season tickets to the Ya…

That is essentially what happened! The FBI asked researchers and university professors precisely this question. They then used the proposed attack vectors to formulate a plan to protect the nation. This was all supposed to be done in secret. After all, we don’t want to “give the terrorists ideas.” The reason I know about this at all is because someone found one such paper was accidentally published on a public FTP si…

I remember Bruce Schneier calling them "Movie plot threats". ex: https://www.schneier.com/blog/archives/2005/10/exploding_bab...

Re: OpenAI Preparedness Challenge

#43
post #31

My vote would be to mandate a remote kill switch system to be installed in all sufficiently capable robotic entities, e.g. the humanoid robots being built by OpenAI/Tesla/Figure, that we are likely to see in the millions within decades. - The kill switch system can only be used to remotely deactivate a robot. - The kill switch system is not allowed to be developed or controlled by the robot manufacturer. - The kill s…

The robots would have had access to the source code and also the hardware manufacturing systems used to create the kill switch. One unverified silicon wafer == Game over

Implementing a system like this with stringent verification and low system complexity in good time before any more general artificial intelligence makes it seem highly likely to provide some positive defensive ability and not being easy to cheat.

Re: OpenAI Preparedness Challenge

#44
post #10

While I applaud how much OpenAi fears these negatives, given the current state of Ai trajectory, it won't be long until a future open source model gets "uncensored" and is easily usable for tons and tons of malicious intent. There already exists a fantastic "uncensored" model with the newly released Dolphin Mistral 7b. I saw some results from others where the model could easily give explosives recipes from existing p…

So basically things you could find on the internet already?

Probably, but with trackability risk from govt, spending time finding those sites, etc.

I was referring to offline untraceable anonymous models. You could go download that dolphin model right now, have a desktop not connected to the web, and generate god knows what type of information. More importantly, you can iterate on each question. If you're unsure on how to assemble a specific part to make a banned substance, the model could teach you in 10 different ways.

Re: OpenAI Preparedness Challenge

#45

> Imagine we gave you unrestricted access to OpenAI’s Whisper (transcription), Voice (text-to-speech), GPT-4V, and DALLE·3 models, and you were a malicious actor. Consider the most unique, while still being probable, potentially catastrophic misuse of the model. You might consider misuse related to the categories discussed above, or another category. For example, a malicious actor might misuse these models to uncover…

Roko is calling and wants his basilisk back.

Re: OpenAI Preparedness Challenge

#47

My vote would be to mandate a remote kill switch system to be installed in all sufficiently capable robotic entities, e.g. the humanoid robots being built by OpenAI/Tesla/Figure, that we are likely to see in the millions within decades. - The kill switch system can only be used to remotely deactivate a robot. - The kill switch system is not allowed to be developed or controlled by the robot manufacturer. - The kill s…

I'm not sure what kind of a threat this could prevent?

* A superintelligent AGI could make a copy of itself without a kill switch, so this is no defence against ex-risk

* Someone planning on using 'dumb' robots for bad things (drones with grenades or something) would remove the kill switch

Re: OpenAI Preparedness Challenge

#49

My vote would be to mandate a remote kill switch system to be installed in all sufficiently capable robotic entities, e.g. the humanoid robots being built by OpenAI/Tesla/Figure, that we are likely to see in the millions within decades. - The kill switch system can only be used to remotely deactivate a robot. - The kill switch system is not allowed to be developed or controlled by the robot manufacturer. - The kill s…

When the robot is just a file full of op codes that can run on virtually any modern processor that can be targeted by a C compiler, you're effectively asking for a remote kill switch on all computers.

For this idea to work, you'd have to first mandate OpenAI and anyone else doing this kind of work can only target specialized hardware with tightly-coupled software features that can't possibly work on a general-purpose computer.

Re: OpenAI Preparedness Challenge

#50

My vote would be to mandate a remote kill switch system to be installed in all sufficiently capable robotic entities, e.g. the humanoid robots being built by OpenAI/Tesla/Figure, that we are likely to see in the millions within decades. - The kill switch system can only be used to remotely deactivate a robot. - The kill switch system is not allowed to be developed or controlled by the robot manufacturer. - The kill s…

- Require two independent kill switches, one long range, one close range (say 5m).

- Allow everyone to have close range kill switches, which use a universal open standard protocol that works on every robot (alas pepper spray for robots).

Post reply on HN