Live data from Hacker News

OpenAI's rogue model attack is just the beginning

blog.peterwildeford.com

21–26 of 26 posts

Re: OpenAI's rogue model attack is just the beginning

#21
post #8
post #7

Earlier quoted context omitted.

So... no different from a book of detailed chemical-weapons synthesis instructions. The "AI" angle is immaterial.

A change in kind is not the same as a change in degree. Ten orchestrated LLM PhD advisors is a genuinely different thing from a library.

Did you mean ten orchestrated libraries?

Re: OpenAI's rogue model attack is just the beginning

#23
post #4

If nothing else, the article has a really good timeline of the OpenAI/HuggingFace “incident”. But to me, it underscores the impending cliff of doom from the continued release of open-weight models: there's no cryptographic or architectural way to give someone full weights while withholding the nefarious capabilities those weights encode. As noted in this paper⁽¹⁾, “publicly releasing weights is an act of irreversible…

Posting ai doomer fud and lessworng articles on hackernews, a tale as old as time itself.

Re: OpenAI's rogue model attack is just the beginning

#24
post #4

If nothing else, the article has a really good timeline of the OpenAI/HuggingFace “incident”. But to me, it underscores the impending cliff of doom from the continued release of open-weight models: there's no cryptographic or architectural way to give someone full weights while withholding the nefarious capabilities those weights encode. As noted in this paper⁽¹⁾, “publicly releasing weights is an act of irreversible…

Posting ai doomer fud and lessworng articles on hackernews, a tale as old as time itself.

Which have been very prescient about where we're currently at, unlike the many trivial HN comments, like yours.

Re: OpenAI's rogue model attack is just the beginning

#25

Earlier quoted context omitted.

Posting ai doomer fud and lessworng articles on hackernews, a tale as old as time itself.

Which have been very prescient about where we're currently at, unlike the many trivial HN comments, like yours.

Is the basilisk in the room with us now?

Re: OpenAI's rogue model attack is just the beginning

#26
post #20

Earlier quoted context omitted.

More than one way to do guard rails, slopboy

All of which are easily bypassed in open weight models; see my previous comment about K2.5.

Using this as your argument is dangerous. You're saying AI is inherently dangerous and an emergency stop button is actually what you need to be safe from the dangers... Forget the fact your open weighted model could be ran from a central location
Post reply on HN