Live data from Hacker News

Anthropic apologizes for invisible Claude Fable guardrails

theverge.com

171–180 of 489 posts

Re: Anthropic apologizes for invisible Claude Fable guardrails

#171

Will Anthropic ever respond to these negative comments here? They won't.

They literally just have. The ethos is explained here. If you don't bother to read or grapple with it that isn't on them.

https://darioamodei.com/post/policy-on-the-ai-exponential

Re: Anthropic apologizes for invisible Claude Fable guardrails

#172

Earlier quoted context omitted.

But the things they say they believe are insane and totally unmoored from physical, societal, and economic reality. If they actually believe those things they're untrustworthy because they're delusional. If they don't, they're untrustworthy because they're fraudulent. Either way it's not good..

They're not. They're in the eye of the storm and see what's going on the clearest. They were ahead of the curve to be where they're at now, and they're still ahead of the curve for where we're going. All the other heads of labs like Sam Altman and Demis have been saying the same thing since 2015-2016 way before any of this "marketing" would ever have been at play.

There's a simpler explanation that fits the data better: they're lying.

Generally, in the past when tech companies have made outlandish claims that were not backed by evidence, they're later found out to have lied. This is an ancient pattern going back to the dotcom era and before, but for recent examples you need only look back a few years to the web3 era. If they're not lying, they can show it by producing the results they claim. Until then, they're probably just lying.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#173

I like Claude Code a lot, I think it sets a dangerous precedent to put guardrails in that return a response from a prompt that was modified by the system in real time in order to subvert the original intent. Fail cleanly. Anything else makes it too difficult to rely on. edit: Giving the absolute maximum benefit of the doubt I understand that they see themselves as "stewards" for lack of a better word. But the EA thin…

That also means people are paying money to execute a prompt they've (partially) written.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#174

Earlier quoted context omitted.

Well, Anthropic thinks it should be the Trump administration [1]. This whole business just keeps getting dumber. 1: https://darioamodei.com/post/policy-on-the-ai-exponential

Read the actual essay. I cannot possibly imagine how you come to that conclusion unless you're just arguing in bad faith.

No. You read the actual essay, then explain how we're supposed to interpret this more charitably:

    Frontier AI models, like airplanes, should 
    be required to go through technical testing 
    and auditing, and their release should be 
    blocked or reversed as a threat to public 
    safety if they do not meet high standards 
    of safety. I am grateful to see the Trump 
    administration’s Executive Order move 
    incrementally towards a greater role for 
    government in AI, though Anthropic’s proposal 
    recommends even further action. 
They are all-but-literally sucking up to the administration that declared their company a supply-chain risk, arguing that the same administration should be given gatekeeping authority over all high-quality LLMs including open-weight releases. Go gaslight somebody else.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#176

Earlier quoted context omitted.

You'd be fine if the PRC gets to ASI first? That's an interesting opinion.

PRC labs reportedly aren't even thinking about getting to ASI, much less trying. They think of AI as a technology that can provide utility across the board even without anything like superhuman smarts.

Nope, they're accelerating towards superhuman smarts as fast as they can too.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#177

Earlier quoted context omitted.

It's not entirely bullshit, but they're continuing to be a terrible company with great products.

you really think they're building anything that's too dangerous for public release though? that's the BS

Honestly, while I love having access to this grade of AI, yeah, it's been too dangerous for a few releases now.

And Fable is cracked. Way better than anything, and the biggest improvements are on the scariest subjects.

So given the state of the world at the moment, and the number of software patches we're barely keeping up with... I'm thankful that they're not making it worse.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#178

Earlier quoted context omitted.

Opus is nowhere close to Fable. Fable feels at least one generation ahead to me. https://x.com/hyperagentapp/status/2064396004032463157 Edit: OpenAI will launch a similar model soon and I can't wait. We are entering a new era of agents.

What does this even mean?

I added a link.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#179
post #83

Earlier quoted context omitted.

"Only I can save us". It's a classic tragedy and cautionary tale. The idea Anthropic was going to speed run AI so they could control the usage and make it "safe" for humanity was never altruistic; it was a HUGE FUCKING RED FLAG.

[flagged]

And? Now all the zero days, if thats true, get discovered and patched instead of being exclusively hoarded by the select few governments and Israeli spyware companies.

Sounds like a great thing to me.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#180

Earlier quoted context omitted.

Opus is nowhere close to Fable. Fable feels at least one generation ahead to me. https://x.com/hyperagentapp/status/2064396004032463157 Edit: OpenAI will launch a similar model soon and I can't wait. We are entering a new era of agents.

Care to share any specifics?

I have a design for a really complex software I want to build and there were gaps I knew of in the design. Opus couldn’t identify them but Fable did. I’m just talking about it reviewing the design, not coding. But yeah, it’s insanely expensive. It does spin off sub agents so I suspect it might be cheaper if you had it create a bunch of plan files and then pointed deepseek at this plan files or something like that
Post reply on HN