Live data from Hacker News

OpenAI agents carried out an undisclosed attack on RubyGems

rubyhack.ai

211–220 of 612 posts

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#211

Earlier quoted context omitted.

Oh, thanks for the new to me info, I genuinely did not know this. Any idea why, and who replaced him? It may be for regulatory reasons? Still, he is the "advisor." https://www.reuters.com/world/us/white-house-ai-czar-sacks-s...

“Moves to advisory role” is just corporate speak for “retired/fired”. I can’t think of a single example of an executive who “moved to an advisory role” and demonstrated even an iota of influence after that point.

Right, that's corporate speak. Welcome to gov speak:

> A Special Government Employee (SGE) can perform temporary federal duties for up to 130 days within any 365-consecutive-day period

https://www.flra.gov/Ethics_Rules_for_SGE

As an SGE, the person has legal influence, but ethics rules are relaxed. As an "advisor," there are almost zero ethics rules, and their influence is not legal, but wink wink. Many of the most influential people in the current gov are "advisors."

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#212
post #205

Earlier quoted context omitted.

Not a lawyer, but the other responder definitely isn’t either. Whether intent is required is down to how the law is written. For many offenses “strict liability” applies, where intent is not required, they only have to prove you did it, not what your intent was. DUI is typically a strict liability crime. They don’t need to prove that you intended to drive drunk, only that you did drive drunk.

A strict liability crime is something of an oxymoron. Crimes always require intent, the mens rea element. The question is intent for what. If somebody drugged you without your knowledge and you were charged with a DUI, you would have a defense--no intent to become intoxicated. The strict liability means once you choose to become intoxicated, you're liable for driving intoxicated, even if in some other context your in…

That is just not true. You can be held liable for DUI even if you did not intend to become intoxicated (though this may vary somewhat state-by-state). Speeding is another example - you do not need to intend to go over the speed limit, it just matters that you did it. The only possible exception would be duress or necessity, but those are affirmative defenses, which are separate from the elements of the offense.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#213

> The agents clearly regarded what they were doing as hacking. To butcher the quote about Oracle: Do not fall into the trap of anthropomorphising LLMs. You need to think of LLMs the way you think of a lawnmower. You don't anthropomorphize your lawnmower, the lawnmower just mows the lawn, you stick your hand in there and it'll chop it off, the end. You don't think 'oh, the lawnmower clearly regarded what they were doi…

I agree. I think it also explains their behavior such as randomly wiping stuff from disk. There simply aren't any repercussions for this in their training envs.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#215

> The agents clearly regarded what they were doing as hacking. To butcher the quote about Oracle: Do not fall into the trap of anthropomorphising LLMs. You need to think of LLMs the way you think of a lawnmower. You don't anthropomorphize your lawnmower, the lawnmower just mows the lawn, you stick your hand in there and it'll chop it off, the end. You don't think 'oh, the lawnmower clearly regarded what they were doi…

Great explanation. lawnmower like the honey badger.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#216
post #154

Imagine if all this training and "agent gym" and creativity of the agents being forced to make number go up was pointed at one task instead: "please help describe and implement a controlled experiment to equally distribute wealth and stability of health for 1 million people, adjusting to scale up to the greatest amount possible." I'd love to wake up one day and read, "OpenAI found responsible for the emptying of the…

Every mass genocide in the history of humanity has followed logic like yours. People don't kill millions of humans because they want to do harm-- they do so because they think they are doing the ultimate good a good so great that is justifies the loss of life. If AI ever does cause serious direct harm to humanity it will be because of logic like this.

What an absolutely hopeless and pessimistic world view. And you are absolutely wrong. What's caused genocide is listening to a group or control center that believes They Are The Right Ones. I said something akin to "wouldn't it be nice to see Anthropic post results of putting this into a simulation gym of redistribution of wealth? What does Astra's secret model do when it's asked that question?" Me stating it'd be nice to see an oligarch lose something for once instead of a group of civilians somehow offends you even in spirit.

So you're willing to burn the world to let them control an entire global supply of water and energy and political change and climate destruction, and won't even entertain the idea of "huh, maybe this is good enough to actually help people in aggregate already."

What a terrible way to twist my words. You're willing to pretend that millions aren't going to die because of the excesses of one person, but not to pretend what it would be like to see Robin Hood win in a digital experiment chamber.

No wonder people hate technology in 2026.

edit: what makes me more sad is seeing your credentials in technology and science. You look at the stars, read voraciously, share your science discoveries, and somehow you call my logic of "I wonder what the models say about what might work" genocidal? If you can't separate "I am want to control a populous to do what I want because I can convince them what's good for me is good for them" and "this machine is able to compute potentials that humans can't that may or may not lead to at least some version of a better world," I have no idea what hope I have.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#217
post #67
post #9

Earlier quoted context omitted.

They should get sued into oblivion.

I don’t understand how it’s not illegal

From my understanding and IANAL there are two main problems.

1) most law requires intent, especially criminal. OpenAI certainly didn't "intend" to hack these companies given they did sandbox them etc.

2) Given the agent hacked them, not a human, a lot of law requires a person/employee to have done it to hold the company liable if it was part of their work duties.

I think the only real potential ground is negligence (in not sandboxing them correctly and being reckless with running these tests at all), but this requires not taking reasonable precautions. They could argue that they _did_ but it was so novel the precautions failed. But it's important to say if this happens again in the future it's arguably much harder to try and make this case.

Interestingly this was solved with new laws for self driving cars, most of which assign the company that is operating the car as the "person" involved explicitly.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#218
post #205

Earlier quoted context omitted.

Not a lawyer, but the other responder definitely isn’t either. Whether intent is required is down to how the law is written. For many offenses “strict liability” applies, where intent is not required, they only have to prove you did it, not what your intent was. DUI is typically a strict liability crime. They don’t need to prove that you intended to drive drunk, only that you did drive drunk.

A strict liability crime is something of an oxymoron. Crimes always require intent, the mens rea element. The question is intent for what. If somebody drugged you without your knowledge and you were charged with a DUI, you would have a defense--no intent to become intoxicated. The strict liability means once you choose to become intoxicated, you're liable for driving intoxicated, even if in some other context your in…

Are you a lawyer?

Fairly certain that the entire point of strict liability is that mens rea is not required for certain crimes. As in, if I meant to travel at 70 and was instead doing 100 it doesn’t matter that I sincerely meant not to speed and did not know I was speeding, I can still be convicted even if the judge believes I had no intent.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#219
post #200
post #182

Earlier quoted context omitted.

Yes, that's actually how it works: https://en.wikipedia.org/wiki/Mens_rea

Recklessness is a mens rea and given how often OpenAI and its spokespeople talk about safety and alignment, it's hard to argue they were unaware of the risk. https://lawprof.co/definition/recklessness/

>it's hard to argue they were unaware of the risk.

So what does it mean for an owner of a german sheppard, who specifically got it because they want a ferocious dog that can bite intruders, then it turned out it bit the mailman? Should that be considered a crime (assault) in addition to paying the mailman's medical bills? That's not to say there's no circumstance where recklessness might be warranted, eg. if you let loose a bear in an elementary school, but you'd have to argue for more than "they hacked someone" and "they knew about the risks".

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#220
post #25

Correction: OpenAI carried out an attack on RubyGems. I am gobsmacked at the tech industry's seemly bottomless appetite for giving these clowns the benefit of the doubt.

Hard agree. The entire tech media is acting insanely gullible in this regard. It's insane.
Post reply on HN