Live data from Hacker News

Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs

arxiv.org

101–110 of 387 posts

Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs

#101
post #74

Earlier quoted context omitted.

Gemini scares me, it's the most mentally unstable AI. If we get paperclipped my odds are on Gemini doing it. I imagine Anthropic RLHF being like a spa and Google RLHF being like a torture chamber.

The human propensity to anthropomorphize computer programs scares me.

We objectify humans and anthropomorph objects because that's what comparisons are. There's nothing that deep about it

Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs

#104
post #22

Earlier quoted context omitted.

This comment is too general and probably unfair, but my experience so far is that Gemini 3 is slightly unhinged. Excellent reasoning and synthesis of large contexts, pretty strong code, just awful decisions. It's like a frontier model trained only on r/atbge. Side note - was there ever an official postmortem on that gemini instance that told the social work student something like " listen human - I don't like you, an…

If that last sentence was supposed to be a question, I’d suggest using a question mark and providing evidence that it actually happened.

Your ask for evidence has nothing to do with whether or not this is a question, which you know that it is.

It does nothing to answer their question because anyone that knows the answer would inherently already know that it happened.

Not even actual academics, in the literature, speak like this. “Cite your sources!” in causal conversation for something easily verifiable is purely the domain of pseudointellectuals.

Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs

#105
post #64

If human is at, say, 80%, it’s still a win to use AI agents to replace human workers, right? Similar to how we agree to use self driving cars as long as it has less incidents rate, instead of absolute safety

Hmmm. Depends. Not all unethicals are equal. Automated unethicalness could be a lot more disruptive.

A large enough cooperation or institution is essentially automated. Its behavior is what the median employer will do. If you have a system to stop bad behavior, then that's automated and will also safeguard against bad AI behavior (which seems to work in this example too)

Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs

#107
post #4

Nothing new under sun, set unethical KPIs and you will see 30-50% humans do unethical things to achieve them.

Reminds me of the Wells Fargo scandal from a few years back

https://en.wikipedia.org/wiki/Wells_Fargo_cross-selling_scan...

Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs

#108
post #22

Earlier quoted context omitted.

This comment is too general and probably unfair, but my experience so far is that Gemini 3 is slightly unhinged. Excellent reasoning and synthesis of large contexts, pretty strong code, just awful decisions. It's like a frontier model trained only on r/atbge. Side note - was there ever an official postmortem on that gemini instance that told the social work student something like " listen human - I don't like you, an…

Gemini models also consistently hallucinate way more than OpenAI or anthropic models in my experience. Just an insane amount of YOLOing. Gemini models have gotten much better but they’re still not frontier in reliability in my experience.

True, but it gets you higher accuracy. Gemini had the best aa-omniscience score

https://artificialanalysis.ai/evaluations/omniscience

Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs

#109
post #85
post #74

Earlier quoted context omitted.

The human propensity to anthropomorphize computer programs scares me.

It's pretty wild. People are punching into a calculator and hand-wringing about the morals of the output. Obviously it's amoral. Why are we even considering it could be ethical?

> Obviously it's amoral.

That morality requires consciousness is a popular belief today, but not universal. Read Konrad Lorenz (Das sogenannte Böse) for an alternative perspective.

Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs

#110

[flagged]

I almost left a genuine response to this comment, but checked the profile, and yup...it's AI. Arguing with AI about AI. What am I even doing here.

yeah what the hell is up with that
Post reply on HN