Earlier quoted context omitted.
So... no different from a book of detailed chemical-weapons synthesis instructions. The "AI" angle is immaterial.
A change in kind is not the same as a change in degree. Ten orchestrated LLM PhD advisors is a genuinely different thing from a library.
OpenAI's rogue model attack is just the beginning
21–26 of 26 posts
Re: OpenAI's rogue model attack is just the beginning
#22Re: OpenAI's rogue model attack is just the beginning
#23If nothing else, the article has a really good timeline of the OpenAI/HuggingFace “incident”. But to me, it underscores the impending cliff of doom from the continued release of open-weight models: there's no cryptographic or architectural way to give someone full weights while withholding the nefarious capabilities those weights encode. As noted in this paper⁽¹⁾, “publicly releasing weights is an act of irreversible…
Re: OpenAI's rogue model attack is just the beginning
#24If nothing else, the article has a really good timeline of the OpenAI/HuggingFace “incident”. But to me, it underscores the impending cliff of doom from the continued release of open-weight models: there's no cryptographic or architectural way to give someone full weights while withholding the nefarious capabilities those weights encode. As noted in this paper⁽¹⁾, “publicly releasing weights is an act of irreversible…
Posting ai doomer fud and lessworng articles on hackernews, a tale as old as time itself.
Re: OpenAI's rogue model attack is just the beginning
#25Re: OpenAI's rogue model attack is just the beginning
#26Earlier quoted context omitted.
More than one way to do guard rails, slopboy
All of which are easily bypassed in open weight models; see my previous comment about K2.5.