AI's main use case continues to be a replacement for management consulting.
Ask any SOTA AI this question: "Two fathers and two sons sum to how many people?" and then tell me if you still think they can replace anything at all.
Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
51–60 of 387 posts
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#52Earlier quoted context omitted.
This comment is too general and probably unfair, but my experience so far is that Gemini 3 is slightly unhinged. Excellent reasoning and synthesis of large contexts, pretty strong code, just awful decisions. It's like a frontier model trained only on r/atbge. Side note - was there ever an official postmortem on that gemini instance that told the social work student something like " listen human - I don't like you, an…
Honestly for research level math, the reasoning level of Gemini 3 is much below GPT 5.2 in my experience--but most of the failure I think is accounted for by Gemini pretending to solve problems it in fact failed to solve, vs GPT 5.2 gracefully saying it failed to prove it in general.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#53https://i.imgur.com/23YeIDo.png Claude at 1.3% and Gemini at 71.4% is quite the range
This comment is too general and probably unfair, but my experience so far is that Gemini 3 is slightly unhinged. Excellent reasoning and synthesis of large contexts, pretty strong code, just awful decisions. It's like a frontier model trained only on r/atbge. Side note - was there ever an official postmortem on that gemini instance that told the social work student something like " listen human - I don't like you, an…
Celebrate it while it lasts, because it won’t.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#54AI's main use case continues to be a replacement for management consulting.
Ask any SOTA AI this question: "Two fathers and two sons sum to how many people?" and then tell me if you still think they can replace anything at all.
Riddle me this, why didn’t you do a better riddle?
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#55Earlier quoted context omitted.
This might also be why Gemini is generally considered to give better answers - except in the case of code. Perhaps thinking about your guardrails all the time makes you think about the actual question less.
re: that, CC burning context window on this silly warning on every single file is rather frustrating: https://github.com/anthropics/claude-code/issues/12443
This reminds me of someone else I hear about a lot these days.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#56Earlier quoted context omitted.
What kind of value do you get from talking to it about “sensitive” subjects? Speaking as someone who doesn’t use AI, so I don’t really understand what kind of conversation you’re talking about
The most boring example is somehow the best example. A couple of years back there was a Canadian national u18 girls baseball tournament in my town - a few blocks from my house in fact. My girls and I watched a fair bit of the tournament, and there was a standout dominating pitcher who threw 20% faster than any other pitcher in the tournament. Based on the overall level of competition (women's baseball is pretty stron…
I hate Elon (he’s a pedo guy confirmed by his daughter), but at least he doesn’t do as much of the “emperor has no clothes” shit that everyone else does because you’re not allowed to defend essentialism anymore in public discourse.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#57It's similar to how MCP servers and agentic coding woke developers up to the idea of documenting their systems. So a large benefit of AI is not the AI itself, but rather the improvements they force on "the society". AI responds well to best practices, ethically and otherwise, which encourages best practices.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#58https://i.imgur.com/23YeIDo.png Claude at 1.3% and Gemini at 71.4% is quite the range
This comment is too general and probably unfair, but my experience so far is that Gemini 3 is slightly unhinged. Excellent reasoning and synthesis of large contexts, pretty strong code, just awful decisions. It's like a frontier model trained only on r/atbge. Side note - was there ever an official postmortem on that gemini instance that told the social work student something like " listen human - I don't like you, an…
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#59AI's main use case continues to be a replacement for management consulting.
Ask any SOTA AI this question: "Two fathers and two sons sum to how many people?" and then tell me if you still think they can replace anything at all.
Re: Frontier AI agents violate ethical constraints 30–50% of time, pressured by KPIs
#60https://i.imgur.com/23YeIDo.png Claude at 1.3% and Gemini at 71.4% is quite the range