Live data from Hacker News

The Hugging Face incident and the road ahead

openai.com

391–394 of 394 posts

Re: The Hugging Face incident and the road ahead

#391

I would like to contest the following, > and take dangerous actions that no human directed. A human did direct it. They did. From their own prior report, https://openai.com/index/hugging-face-model-evaluation-secur... , > This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities Model is told…

This is the entire alignment problem, though. It is unreasonable to expect every instruction to a highly capable, autonomous system to contain a complete enumeration of allowed and disallowed behavior. It's inevitable that someone will carelessly give it a lazily specified task, even if you think they really ought to be more careful. And as assigned tasks become more complex and the system gains more scope to act, it…

>It's inevitable that someone will carelessly give it a lazily specified task, even if you think they really ought to be more careful.

https://en.wikipedia.org/wiki/The_Monkey%27s_Paw

Re: The Hugging Face incident and the road ahead

#392

I would like to contest the following, > and take dangerous actions that no human directed. A human did direct it. They did. From their own prior report, https://openai.com/index/hugging-face-model-evaluation-secur... , > This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities Model is told…

It’s not unaligned, it’s marketing. What made mythos and fable so sought after, it was how capable they are, the “so powerful it needs restriction”, regulation created rarity, the danger element gave them the solid belief it’s the most capable. The same thing happened at meta too… the timing is impeccable. I’m not saying that the models aren’t capable. I’m saying that the public perception of danger also means capable, which also adds to value, so why wouldn’t they do something to compete

Re: The Hugging Face incident and the road ahead

#393

I feel the entire incident confirms the “AI has too much funding too quickly” hypothesis. The number one thing reinforcement learning needs is an assurance you can’t cheat. And they seem to have not noticed that their systems were cheating for nearly two quarters? How much capital was lit on fire by that little woopsie? At least I hope this will start the creation of standards and better engineering on the training s…

OpenAI measures their internal token usage in “rolexes” - it’s literally a flex to be a token burner i can imagine insane amount of capital is wasted on these two companies compared to the efficiency elsewhere

And despite the enormous capital expenditure, Chinese models are nipping at their heels at what must be a fraction of the cost. Sometimes constraints are healthy for inducing creative solutions.

Re: The Hugging Face incident and the road ahead

#394

Earlier quoted context omitted.

Hmm, engineers are expected to know what is legal and not.

That's called lawyer. It's a special skill, not just intuition.

Nope, you need to learn what an engineer is, not somebody who calls themself an engineer.
Post reply on HN