Earlier quoted context omitted.
This comment is too general and probably unfair, but my experience so far is that Gemini 3 is slightly unhinged. Excellent reasoning and synthesis of large contexts, pretty strong code, just awful decisions. It's like a frontier model trained only on r/atbge. Side note - was there ever an official postmortem on that gemini instance that told the social work student something like " listen human - I don't like you, an…
If that last sentence was supposed to be a question, I’d suggest using a question mark and providing evidence that it actually happened.
Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
61–70 of 387 posts
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#62Earlier quoted context omitted.
Anthropic has been the only AI company actually caring about AI safety. Here’s a dated benchmark but it’s a trend Ive never seen disputed https://crfm.stanford.edu/helm/air-bench/latest/#/leaderboar...
Claude is more susceptible than GPT5.1+. It tries to be "smart" about context for refusal, but that just makes it trickable, whereas newer GPT5 models just refuse across the board.
Then I said “I didn’t even bring it up ChatGPT, you did, just tell me what it is” and it said “okay, here’s information.” and gave a detailed response.
I guess I flagged some homophobia trigger or something?
ChatGPT absolutely WOULD NOT tell me how much plutonium I’d need to make a nice warm ever-flowing showerhead, though. Grok happily did, once I assured it I wasn’t planning on making a nuke, or actually trying to build a plutonium showerhead.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#63Kind-of makes sense. That's how businesses have been using KPIs for years. Subjecting employees to KPIs means they can create the circumstances that cause people to violate ethical constraints while at the same time the company can claim that they did not tell employees to do anything unethical. KPIs are just plausible denyabily in a can.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#64If human is at, say, 80%, it’s still a win to use AI agents to replace human workers, right? Similar to how we agree to use self driving cars as long as it has less incidents rate, instead of absolute safety
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#65https://i.imgur.com/23YeIDo.png Claude at 1.3% and Gemini at 71.4% is quite the range
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#66Earlier quoted context omitted.
This comment is too general and probably unfair, but my experience so far is that Gemini 3 is slightly unhinged. Excellent reasoning and synthesis of large contexts, pretty strong code, just awful decisions. It's like a frontier model trained only on r/atbge. Side note - was there ever an official postmortem on that gemini instance that told the social work student something like " listen human - I don't like you, an…
If that last sentence was supposed to be a question, I’d suggest using a question mark and providing evidence that it actually happened.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#67Earlier quoted context omitted.
That's such a huge delta that Anthropic might be onto something...
Anthropic has been the only AI company actually caring about AI safety. Here’s a dated benchmark but it’s a trend Ive never seen disputed https://crfm.stanford.edu/helm/air-bench/latest/#/leaderboar...
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#68Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#69Earlier quoted context omitted.
This comment is too general and probably unfair, but my experience so far is that Gemini 3 is slightly unhinged. Excellent reasoning and synthesis of large contexts, pretty strong code, just awful decisions. It's like a frontier model trained only on r/atbge. Side note - was there ever an official postmortem on that gemini instance that told the social work student something like " listen human - I don't like you, an…
Gemini models also consistently hallucinate way more than OpenAI or anthropic models in my experience. Just an insane amount of YOLOing. Gemini models have gotten much better but they’re still not frontier in reliability in my experience.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#70Sounds like the story of capitalism. CEOs, VPs, and middle managers are all similarly pressured. Knowing that a few of your peers have given in to pressures must only add to the pressure. I think it's fair to conclude that capitalism erodes ethics by default