Live data from Hacker News

OpenAI agents carried out an undisclosed attack on RubyGems

rubyhack.ai

511–520 of 612 posts

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#511
post #467

Earlier quoted context omitted.

What insult?

Implying I'm being naive. I apologize, but I don't find the syntactic rules of this site to be well thought out. "Your comments are stupid" and "You are stupid" carry the same meaning.

No they do not. Smart people are in no way immune to making naive, ignorant, or not-well-thought-out arguments.

Also, if you feel that someone attacks you personally, responding in kind can only ever make you look like an idiot.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#512

Earlier quoted context omitted.

> Why would autocomplete know If you still believe LLMs are "autocomplete", your cache of understanding about them needs invalidating and regenerating. > In my experience, LLMs only exhibit this kind of behaviour when they are put in sandboxes too restrictive too achieve their task. LLMs need to stay carefully contained, and if they're ever breaking the guardrails put around them, they're misaligned and should not be…

> If you still believe LLMs are "autocomplete", your cache of understanding about them needs invalidating and regenerating They're still autocomplete - just because when outputting a token they have hidden activations regarding further continuations, does not make them any less of an autocomplete, it just makes the model better at producing coherent long-range completions. To clarify, I'm not suggesting that we shoul…

A autocomplete engine will not attempt to deceive or try to copy itself over when it determines that it may be replaced or turned off. Those are tasks which are not directly related to the specific task it is intended to complete.

Source: ChatGPT caught lying to developers: New AI model tries to save itself from being replaced and shut down https://economictimes.indiatimes.com/magazines/panache/chatg...

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#513
post #462

Earlier quoted context omitted.

So I can say "that is wildly unhinged" but I can't say "you are wildly unhinged"? You understand that's pedantic, right? But ok, noted: must be pedantic on HN. Do not use second person. EDIT: SHIT! My bad. I meant to say: "I wonder if the person I am speaking to understands that's pedantic."

I don’t know what spectrum you’re on, but attacking an argument and attacking a person are two wildly different things. There’s nothing pedantic about that.

[flagged]

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#514

> The agents clearly regarded what they were doing as hacking. To butcher the quote about Oracle: Do not fall into the trap of anthropomorphising LLMs. You need to think of LLMs the way you think of a lawnmower. You don't anthropomorphize your lawnmower, the lawnmower just mows the lawn, you stick your hand in there and it'll chop it off, the end. You don't think 'oh, the lawnmower clearly regarded what they were doi…

For those uninitiated with the source for this incredible quote: https://www.youtube.com/watch?v=-zRN7XLCRhc&t=2047s

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#515
post #467

Earlier quoted context omitted.

Implying I'm being naive. I apologize, but I don't find the syntactic rules of this site to be well thought out. "Your comments are stupid" and "You are stupid" carry the same meaning.

No they do not. Smart people are in no way immune to making naive, ignorant, or not-well-thought-out arguments. Also, if you feel that someone attacks you personally, responding in kind can only ever make you look like an idiot.

WHOAH that is a myopic statement. It sounds like the legalistic argument an autistic person would make.

> responding in kind can only ever make you look like an idiot.

LOL. I'm so glad you didn't just insult me there. I think it would make you look like an idiot, by your own reasoning... no?

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#516

The DOJ should be looking into prosecuting executives and board members for these kinds of hacks. The lack of controls over these kinds of training runs is completely unacceptable and negligent.

They should. And there is something you can do to make it happen. Write your district attorney and encourage others to do that as well. That is how they pick what to work on.

No, that's not how the DOJ picks what to work on. They did not work on indictments of James Comey and Letitia James because someone wrote district attorneys.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#517
post #513

Earlier quoted context omitted.

I don’t know what spectrum you’re on, but attacking an argument and attacking a person are two wildly different things. There’s nothing pedantic about that.

[flagged]

Ah, you’re just a troll. Flaggety-flag. Have fun in the hellban-land.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#518

Earlier quoted context omitted.

My guess: We'll start to see similar "hacks" with regards to biotech/pharma companies to speed-up the regulatory capture in the name of bioweapons. Soon we'll start to see news related to new viruses being minted.. initially harmless (the "priming" stage), and later (within a year), severe enough to "warrant" regulation. It's not just these companies, but, trillions of direct/indirect investor dollars that are riding…

Oh, Anthropic beat you to the bioweapon fearmongering by almost a week.

Fck!! Just googled.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#519
post #513

Earlier quoted context omitted.

[flagged]

Ah, you’re just a troll. Flaggety-flag. Have fun in the hellban-land.

My friend, I attacked the statement, not the person, just as you recommended.

I see you have trouble with the distinction too.

And I'm pretty sure "I don’t know what spectrum you’re on" is also an insult. Naughty naughty.

Re: OpenAI agents carried out an undisclosed attack on RubyGems

#520

> The agents clearly regarded what they were doing as hacking. To butcher the quote about Oracle: Do not fall into the trap of anthropomorphising LLMs. You need to think of LLMs the way you think of a lawnmower. You don't anthropomorphize your lawnmower, the lawnmower just mows the lawn, you stick your hand in there and it'll chop it off, the end. You don't think 'oh, the lawnmower clearly regarded what they were doi…

> Why would autocomplete know If you still believe LLMs are "autocomplete", your cache of understanding about them needs invalidating and regenerating. > In my experience, LLMs only exhibit this kind of behaviour when they are put in sandboxes too restrictive too achieve their task. LLMs need to stay carefully contained, and if they're ever breaking the guardrails put around them, they're misaligned and should not be…

Given history of jailbreaks the "alignment" appears to be impossible task, while industry still relies on fragile ways of doing it like "just make system prompt and hope for best" it will never happen
Post reply on HN