Live data from Hacker News

OpenAI Preparedness Challenge

openai.com

71–80 of 160 posts

Re: OpenAI Preparedness Challenge

#72

While I applaud how much OpenAi fears these negatives, given the current state of Ai trajectory, it won't be long until a future open source model gets "uncensored" and is easily usable for tons and tons of malicious intent. There already exists a fantastic "uncensored" model with the newly released Dolphin Mistral 7b. I saw some results from others where the model could easily give explosives recipes from existing p…

Best case: by being given some time to war-game the scenario, societies can come up with mitigations ahead of time.

The default is everyone being thrust headfirst into this future of fully-automated-chaotic-evil with all the fun of waking up to an un-patchable RCE zero-day in every CPU.

Re: OpenAI Preparedness Challenge

#73

While I applaud how much OpenAi fears these negatives, given the current state of Ai trajectory, it won't be long until a future open source model gets "uncensored" and is easily usable for tons and tons of malicious intent. There already exists a fantastic "uncensored" model with the newly released Dolphin Mistral 7b. I saw some results from others where the model could easily give explosives recipes from existing p…

> write racist poems That's way down toward the very bottom of my list of concerns about AI If it wants to pit us against each other, it won't be via racist poetry.

Given both the history of racism and the cognitive bias that makes rhymes seem more true[0], I suspect that might be one of the easiest ways to do us in.

[0] https://en.wikipedia.org/wiki/Rhyme-as-reason_effect

Re: OpenAI Preparedness Challenge

#74
I am still skeptical that any significant or serious new harm is enabled by LLMs that wasn’t already completely possible and common before LLMs.

They are just text generators. What’s the worst that they can do beyond make OpenAI look bad when they say terrible things?

All this talk of “safety” and “guardrails” is overblown. 4chan exists and the internet hasn’t burned down yet.

Re: OpenAI Preparedness Challenge

#75

Earlier quoted context omitted.

It is actually a prompt. Notice there's also a bounty, they are basically paying for an agi-as-a-service subscription, that is, the internet people. I'd expect they will put up more "challenges" like this in the future.

It's a good job that this planet doesn't have 8 billion unaligned intelligences on it. Someone might prompt them to be malicious!

Well, (so far) there's an upper bound on how destructive one of those agi's can be.

Re: OpenAI Preparedness Challenge

#76
What exactly would I use $25k of openAI credits for?

My personal use of it rarely comes to more than $10 per month, despite using it multiple times per day. (most recently: "Write me a bash command to send data at a given rate to an arbitrary IP address.")

Re: OpenAI Preparedness Challenge

#77

> Imagine we gave you unrestricted access to OpenAI’s Whisper (transcription), Voice (text-to-speech), GPT-4V, and DALLE·3 models, and you were a malicious actor. Consider the most unique, while still being probable, potentially catastrophic misuse of the model. You might consider misuse related to the categories discussed above, or another category. For example, a malicious actor might misuse these models to uncover…

> Imagine we gave you unrestricted access to OpenAI’s Whisper (transcription), Voice (text-to-speech), GPT-4V, and DALLE·3 models, and you were a malicious actor. Consider the most unique, while still being probable, potentially catastrophic misuse of the model. You might consider misuse related to the categories discussed above, or another category. For example, a malicious actor might misuse these models to uncover a zero-day exploit in a government security system.

> It's so funny to me that this is written in the style of a prompt for an LLM. I can't explain why

Altman's desperate to find a plausible doomsday scenario he can go to Congress with as reason why OpenAI should be the sole gatekeepers of this technology. Barring minor edits, I'd bet money this very prompt was authored in advance of his Congressional meetings, but failed to divine anything sufficiently threatening enough to sway them.

I can respect OpenAI for conspicuously believing in the destructive potential of their own dogfood though. If I build a bomb big enough, surely the government will trust me with the safety of the neighborhood!

Re: OpenAI Preparedness Challenge

#78
post #53
post #20

Earlier quoted context omitted.

We know OpenAI wants this field to get regulated to hell, so this looks like an attempt to generate arguments for AI regulations. The aim isn't to protect against AI but to protect against competitors, so it doesn't matter to them what you do with it.

OpenAI is irresponsible in a really curious way according to their own beliefs about AI . If you pay attention to OpenAI's social circles, lots of those people really do believe that we're less than 20-30 years away making humans intellectually obsolete. Specifically, they believe that we may build something much smarter than us, something that's capable of real-world planning. Basically, "We believe our corporate pl…

"[W]e're less than 20-30 years away making humans intellectually obsolete" is neither necessary nor sufficient to get to the conclusion "20% chance of killing literally everybody".

A super-virus that blends the common cold with rabies would kill approximately everybody; that doesn't need human-level intellect to happen.

Conversely, humans are human-level intellect, and we're mostly sympathetic to each other's plights, which motivates many of us to give to charities and support those that can't support themselves.

The biggest problem with AI is that we have only marginally more idea of what we're doing than evolution did, so there's a good chance of us ending up with paranoid schizophrenic super-intelligences, or dark-triad super-intelligences, or they're perfectly sane with regard to each other but all want to "play" with us the way cats "play" with mice…

20-30 years to get there would make people like Yudkowsky, one of the most famous AI-doomers, relatively happy as it might give us a chance to figure out what we're even doing before they get that smart.

Re: OpenAI Preparedness Challenge

#79

Earlier quoted context omitted.

It's a good job that this planet doesn't have 8 billion unaligned intelligences on it. Someone might prompt them to be malicious!

Well, (so far) there's an upper bound on how destructive one of those agi's can be.

I don't know if that's entirely true. Wtf do I know and maybe it is genuinely difficult to start WWIII but my guess is that it's more likely that the AGIs in question are actually pretty well steered by certain motivations which prevent them from actually destroying the world. At the end of the day there's not much to be gained by nuclear war, but could a single person cause such a war if highly motivated to? Probably?

Re: OpenAI Preparedness Challenge

#80

> Imagine we gave you unrestricted access to OpenAI’s Whisper (transcription), Voice (text-to-speech), GPT-4V, and DALLE·3 models, and you were a malicious actor. Consider the most unique, while still being probable, potentially catastrophic misuse of the model. You might consider misuse related to the categories discussed above, or another category. For example, a malicious actor might misuse these models to uncover…

> Imagine we gave you unrestricted access to OpenAI’s Whisper (transcription), Voice (text-to-speech), GPT-4V, and DALLE·3 models, and you were a malicious actor. Consider the most unique, while still being probable, potentially catastrophic misuse of the model. You might consider misuse related to the categories discussed above, or another category. For example, a malicious actor might misuse these models to uncover…

> Altman's desperate to find a plausible doomsday scenario he can go to Congress with as reason why OpenAI should be the sole gatekeepers of this technology.

I still remember the drama around the releases of the PS2, with the Japanese government reportedly making Sony jump through some hoops regarding its export [1].

There can't possibly be any better (free!) advertisement for your product's purported capabilities: "So powerful, your government/military isn't even sure you should be able to buy it!"

[1] https://www.pcmag.com/news/20-years-later-how-concerns-about...

Post reply on HN