Live data from Hacker News

Anthropic apologizes for invisible Claude Fable guardrails

theverge.com

361–370 of 489 posts

Re: Anthropic apologizes for invisible Claude Fable guardrails

#361

Earlier quoted context omitted.

This logic works only if distilling Claude is the only way to create another SOTA LLM, which is not the case.

How do you think the Qwen and MiniMax models perform so similarly to Anthropic frontier models? What is your take then?

They probably stole all the same copyrighted IP

Re: Anthropic apologizes for invisible Claude Fable guardrails

#362

Earlier quoted context omitted.

Even if you believe the concerns have merit, it's hard not to be cynical about people (e.g. Anthropic leadership) paying lip service to those concerns while so obviously leveraging their power and wealth (which depend, by the way, on accelerating the world toward those hypothetical "concerning" scenarios as fast as possible ) to position themselves such that they will become unimaginably rich er if things go their wa…

> (which depend, by the way, on accelerating the world toward those hypothetical "concerning" scenarios as fast as possible) Yes, this dynamic is exactly the one that anyone who's concerned about AI is concerned about. I don't know why you state this as if it's evidence against the concerns lol. Someone being concerned about the incentives of a situation doesn't de facto make them immune to those incentives, obviousl…

This is an excellent comment, and I agree. I do think that there’s also evidence that Altman’s behavior can also be explained as a person who is naturally manipulative also being stuck in the trap and responding to incentives. But not necessarily a snake just in it for himself. The thing I keep coming back to about Altman: he doesn’t have any equity in OpenAI. And he definitely could have if he’d wanted. It’s hard for me to square that with the idea of him being greedy and self-interested.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#363

Earlier quoted context omitted.

No. You read the actual essay, then explain how we're supposed to interpret this more charitably: Frontier AI models, like airplanes, should be required to go through technical testing and auditing, and their release should be blocked or reversed as a threat to public safety if they do not meet high standards of safety. I am grateful to see the Trump administration’s Executive Order move incrementally towards a great…

You got baited by a confirmed Anthropic shill, see more info here: https://news.ycombinator.com/item?id=48270186

Wait, did you actually claim that most work at FAANGs doesn’t require an NDA and that was evidence to support your accusation?

Re: Anthropic apologizes for invisible Claude Fable guardrails

#365
post #312

Earlier quoted context omitted.

How do you think the Qwen and MiniMax models perform so similarly to Anthropic frontier models? What is your take then?

Probably the same reason a Epyc 9965 from hetzner performs just as well as one from AWS for one tenth the cost. Anthropic is offering a commodity product and trying to convince you it isn’t. It’s even in the name, it’s a myth and a fable. Never happened doesn’t exist. Also I believe at least on coding that qwen is now the frontier model, fable is its copy of frontier models. In the same way that the Ferrari Luce is a…

> Also I believe at least on coding that qwen is now the frontier model

The delusions people live in just to be a hater.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#366
post #352

Earlier quoted context omitted.

>be me >anthropic > mine the internet for data, blasting millions of blogs with scrapers >a few have to shut down, but that's just the price to pay >finally, the chatbot is ready >learn that there are EVIL cretins out there trying to scrape automated output from OUR product to build their chatbot >build in safeguards to new model to stop this >the users are mad, now the model accuses users of being bioterrorists if t…

Seriously... the gaul of people just scraping a model for free data!

You wouldn't download an LLM for free, would you?

Re: Anthropic apologizes for invisible Claude Fable guardrails

#367

Earlier quoted context omitted.

> guard against other people building ASI before they do because they think they are uniquely safety oriented relative to their competitors All this longtermism though is harmful. There are real problems of data theft, bias, labor displacement, and environmental costs that are happening right now but every push for regulation and regulatory capture, and all the safety talk, is always focused on some speculative futur…

I cannot overstate how much I think this take is wrong. Please please reconsider, look at the rate of progress being made, and consider that even if you only think ASI 'may' never happen in your lifetime it should still be one of your #1 concerns. Honestly, that respect for 'copyright protections' has somehow become a leftist shibboleth is bizarre to me and indicative that something has become deeply warped in our di…

it shouldnt.

power consumption and global climate change should

ASI should be in the top 10k concerns maybe, but way below what to eat for dinner.

much higher on the fears is some hype guy pretending he has made this thing, and giving it access to too much stuff, which it then randomly deletes or misuses

it should also be in thes same range as "what if the dinosaurs came back and ate everyone"

theres tons of progress on that too. same with finding aliens

there are real present concerns to worry about, like genocides, concentration camps for immigrants, food costs next winter, ongoing wars in the middle east and europe, etc

all kinds of actually pressing stuff, that doesnt first require burning a couple trillion dollars and forcing poor people to pay through the teeth for their electricity

Re: Anthropic apologizes for invisible Claude Fable guardrails

#368

Earlier quoted context omitted.

There's nothing warped about it at all. Like it or not, it is a real issue. It's also an issue of license washing GPL code to privatize it. It's full scale theft of collective human knowledge, being sold back to us in a for profit private product. Outside of that though, there are other issues right now that need addressed before we speculate about what might be possible with ASI in the future. If the potential for a…

The "global stop order" is just generally perceived as an impossible coordination problem. So instead we see a mix of labs voluntarily putting in guardrails and regulatory efforts (which are not only aimed at hypothetical super-AIs of the future). Of course labs are also in a competitive race. And I actually think that it does make sense that the richest companies in the most dominant positions would in a better posi…

> the capabilities we will have two years from now are hard to even imagine properly.

unless the bitter pill is gone, extraordinarily not this. The capabilities will be limited by the training data we can create to pull information and patterns from

and then we will still be limited by compute, space, and power

mass devaluing of labour isnt particularly believable when everyones predicting that all the big labs are gonna go under trying to subsidize tokens.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#369

Earlier quoted context omitted.

I think it's also worth noting that EA is closely linked to utilitarianism. Most of the pitfalls that people see in EA are the same pitfalls that are classic to utilitarianism, a la "we're going to do this thing we know is locally-bad, because we have a lot of confidence in other effects that are universally-good".

EA essentially just is utilitarianism + a specific type of culture/community.

not to mention all the theft and feeling good about yourself being rich

Re: Anthropic apologizes for invisible Claude Fable guardrails

#370
post #353

Earlier quoted context omitted.

Have you copied Windows and tried to run it? I would love to see the plain text source code that you claim to have. We all would.

half of the developing world did. guess what it stopped a bit the trend? protection.

Did it really? Here in my at least, afaik no one's stopped pirating. The tools to activate may have changed but haven't gone away.
Post reply on HN