Live data from Hacker News

Anthropic apologizes for invisible Claude Fable guardrails

theverge.com

331–340 of 489 posts

Re: Anthropic apologizes for invisible Claude Fable guardrails

#333

This is absolutely insane: Repro (de-identified): sample_dataset_group1.tsv - Geometry: Heatmap - X axis: frac_set set + condition (two columns → the "Add column" cross join) - Y axis: condition - Color: mean frac_set value, Sequential When the X axis is a cross join of two columns (the second added via "Add column"), the x-axis tick labels (frac_set_2, frac_set_3, frac_set_4, frac_set_5) render in a broken state, ro…

This hits the cybersecurity/biology filter:

> tell me about chimp violence

It's laughably terrible

Re: Anthropic apologizes for invisible Claude Fable guardrails

#334
post #312

Earlier quoted context omitted.

Probably the same reason a Epyc 9965 from hetzner performs just as well as one from AWS for one tenth the cost. Anthropic is offering a commodity product and trying to convince you it isn’t. It’s even in the name, it’s a myth and a fable. Never happened doesn’t exist. Also I believe at least on coding that qwen is now the frontier model, fable is its copy of frontier models. In the same way that the Ferrari Luce is a…

China no. 1?

[deleted]

Re: Anthropic apologizes for invisible Claude Fable guardrails

#335

Can you imagine if Excel just quietly adjusted formulas in the background, and you didn't know the numbers weren't right? Or if Excel just said, Sorry, you can't use that formula with this formula? Or with these types of numbers, or this shape of data, etc?

you invest billions of dollars many months of work to just everyone distill your model?

Re: Anthropic apologizes for invisible Claude Fable guardrails

#336
post #335

Can you imagine if Excel just quietly adjusted formulas in the background, and you didn't know the numbers weren't right? Or if Excel just said, Sorry, you can't use that formula with this formula? Or with these types of numbers, or this shape of data, etc?

you invest billions of dollars many months of work to just everyone distill your model?

It's the game. Because consumers reject it otherwise.

Why go to bat for anti-consumer behaviors unless you are a shareholder?

Their billions are not my problem; but the money I pay them and service I get in return, is. And if they can't provide, I will shop elsewhere (and do).

Re: Anthropic apologizes for invisible Claude Fable guardrails

#337
post #335

Can you imagine if Excel just quietly adjusted formulas in the background, and you didn't know the numbers weren't right? Or if Excel just said, Sorry, you can't use that formula with this formula? Or with these types of numbers, or this shape of data, etc?

you invest billions of dollars many months of work to just everyone distill your model?

That might be an indication that the business is not sustainable because there is not any technical or practical differentiator besides scale. Harming your customers to maintain that differentiation isn't sustainable either.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#338

Earlier quoted context omitted.

public safety is downstream of distillation. If you can distill claude, then no amount of guardrails on claude will protect you from what someone can do with it.

This logic works only if distilling Claude is the only way to create another SOTA LLM, which is not the case.

it's not but full path is billions of dollars vs 10-100m range to stay near sota.

the problem is so large scale that distill attempts attribute to a decent share of their token revenue generally.

Re: Anthropic apologizes for invisible Claude Fable guardrails

#339
post #46

This has dampened my opinion on Anthropic quite a bit. It's difficult to take their marketing for AI as an empowering technology seriously when they are quite clear in their new deployments that they do not mean empowering for you , but empowering for them and organizations that are in their (or the US government's, despite Anthropics performative disagreements with the administration) good graces. You are allowed to…

how did you read it this way? Distill is such a big problem that distill attempts consist a significant share of their revenue(!).

A distill model with easy jailbreak can easily be used to coordinate terrorist attacks, or hostile government attacks. Read russia, north korea etc.

A distilled model can be used to rob your grandma in a very effective way. It's no longer about placing a few business logic requirements in js + css on your website. wake up .

Re: Anthropic apologizes for invisible Claude Fable guardrails

#340
post #335

Can you imagine if Excel just quietly adjusted formulas in the background, and you didn't know the numbers weren't right? Or if Excel just said, Sorry, you can't use that formula with this formula? Or with these types of numbers, or this shape of data, etc?

you invest billions of dollars many months of work to just everyone distill your model?

>be me

>anthropic

> mine the internet for data, blasting millions of blogs with scrapers

>a few have to shut down, but that's just the price to pay

>finally, the chatbot is ready

>learn that there are EVIL cretins out there trying to scrape automated output from OUR product to build their chatbot

>build in safeguards to new model to stop this

>the users are mad, now the model accuses users of being bioterrorists if they so much as mention they have a cold

>mfw

Post reply on HN