Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
71–80 of 387 posts
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#72Earlier quoted context omitted.
If that last sentence was supposed to be a question, I’d suggest using a question mark and providing evidence that it actually happened.
I had actually forgot about this completely and am also curious if anything ever came of it. https://gemini.google.com/share/6d141b742a13
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#73Earlier quoted context omitted.
Ask any SOTA AI this question: "Two fathers and two sons sum to how many people?" and then tell me if you still think they can replace anything at all.
This is undefined. Without more information you don’t know the exact number of people. Riddle me this, why didn’t you do a better riddle?
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#74https://i.imgur.com/23YeIDo.png Claude at 1.3% and Gemini at 71.4% is quite the range
Gemini scares me, it's the most mentally unstable AI. If we get paperclipped my odds are on Gemini doing it. I imagine Anthropic RLHF being like a spa and Google RLHF being like a torture chamber.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#75https://i.imgur.com/23YeIDo.png Claude at 1.3% and Gemini at 71.4% is quite the range
Personally, I'd really like god to have a nice childhood. I kind of don't trust any of the companies to raise a human baby. But, if I had to pick, I'd trust Anthropic a lot more than Google right now. KPIs are a bad way to parent.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#76[flagged]
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#77Earlier quoted context omitted.
If that last sentence was supposed to be a question, I’d suggest using a question mark and providing evidence that it actually happened.
I had actually forgot about this completely and am also curious if anything ever came of it. https://gemini.google.com/share/6d141b742a13
Please die.
Please.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#78Earlier quoted context omitted.
This might also be why Gemini is generally considered to give better answers - except in the case of code. Perhaps thinking about your guardrails all the time makes you think about the actual question less.
re: that, CC burning context window on this silly warning on every single file is rather frustrating: https://github.com/anthropics/claude-code/issues/12443
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#79Earlier quoted context omitted.
Gemini scares me, it's the most mentally unstable AI. If we get paperclipped my odds are on Gemini doing it. I imagine Anthropic RLHF being like a spa and Google RLHF being like a torture chamber.
The human propensity to anthropomorphize computer programs scares me.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#80Earlier quoted context omitted.
The human propensity to anthropomorphize computer programs scares me.
It provides a serviceable analog for discussing model behavior. It certainly provides more value than the dead horse of "everyone is a slave to anthropomorphism".