Live data from Hacker News

The classifiers Anthropic puts in front of Fable are too zealous

combine-lab.github.io

121–130 of 203 posts

Re: The classifiers Anthropic puts in front of Fable are too zealous

#121
I totally agree. I have been a Claude fanboy for a while now, but Fable woke me up, and I am currently looking for alternatives.

I don't care how capable it is, if it's going to treat me like it's babysitting a terrorist, it can eff off.

Plain and simple.

Re: The classifiers Anthropic puts in front of Fable are too zealous

#122
post #5

This post can essentially be distilled down to: yes, Fable's classifier (which is meant to downgrade cybersecurity, biology, or jailbreak attempts to Opus 4.8) is definitely overly sensitive to the point of uselessness. e.g. a colleague asked Fable to help create an simple app to help calculate the statistics for phase II and III trials. (Ignoring that such things already exist) it passed his request down to Opus, de…

I'm working on cryptography, all from academic research papers. Started well, but it eventually got some word into its context that is on the banlist. I found that if you tell it to fire off clean Fable subagents and you instruct it to check the Claude Code billing data to check for downgrades, you can get most high-sensitivity spec/review tasks done with Fable. Most. I figure that once GPT 5.6 comes out, Anthropic w…

I have been using GLM because of this reason. I think whoever makes model ignore stupid safety thing is going to win in long run.

Re: The classifiers Anthropic puts in front of Fable are too zealous

#124

To summarize: the classifiers Anthropic puts in front of Fable are way, way too zealous and have way too many false positives. From my experience, the model itself is very useful when it isn't refusing any of your prompts.

Thanks - that's a good phrase we can use to replace the baity title. I've done so above.

(Normally we prefer to find a representative phrase from the article itself, but I found that too daunting and gave up.)

Re: The classifiers Anthropic puts in front of Fable are too zealous

#125
post #81
post #5

This post can essentially be distilled down to: yes, Fable's classifier (which is meant to downgrade cybersecurity, biology, or jailbreak attempts to Opus 4.8) is definitely overly sensitive to the point of uselessness. e.g. a colleague asked Fable to help create an simple app to help calculate the statistics for phase II and III trials. (Ignoring that such things already exist) it passed his request down to Opus, de…

I've been working with it heavily since its first release. I use it for software architects, complex debugging and some development and I have not had it refuse or downgrade even once.

I got downgraded for the first time today. Because I was using a library with the characters "bio" in it. The classifier is strict beyond reason. It got the name from a commit message in the git history (wasn't even in my prompt) and it immediately freaked out. I eventually got it to work by getting opus to write a plan, then editing the plan to strip out all references including commit hashes, then getting Fable to review and refine that edited plan. Eventually got it done. But what a pain.

That said, I've got it easy. My colleagues who are chemists and biologists can't even ask one question. There are so many triggers in their memories and workspaces they can't even ask a non-triggering question. And we all work in medical diagnostics, it's not like we're doing anything remotely nefarious. Fable could be such a benefit, but the limitations make it worthless.

Re: The classifiers Anthropic puts in front of Fable are too zealous

#126
post #69
post #5

This post can essentially be distilled down to: yes, Fable's classifier (which is meant to downgrade cybersecurity, biology, or jailbreak attempts to Opus 4.8) is definitely overly sensitive to the point of uselessness. e.g. a colleague asked Fable to help create an simple app to help calculate the statistics for phase II and III trials. (Ignoring that such things already exist) it passed his request down to Opus, de…

My experience, too. I work on nothing in any way related to cybersecurity or biology. I asked it a few purely mathematical questions, it refused immediately. Before the export embargo I did get it to look at some hairy problems and the output was genuinely useful...

Were your question in math areas related to ML? They also restrict model development and research pretty heavily.

Re: The classifiers Anthropic puts in front of Fable are too zealous

#127
post #43

[flagged]

I havent used fable, but does it cost you the same when it downgrades to opus?

I don't think they charge if you downgrade, but if you upgrade from opus to fable they will charge you.

Re: The classifiers Anthropic puts in front of Fable are too zealous

#128
I am wondering about the author's allegation that there is a user filter, not just a prompt filter.

Of course it could also be the case that it is just a prompt filter, but Fable sees memories from the authors' prior sessions that cause a rejection. I wonder if the author could control for this is in some way, if Claude lets you run isolated session without memory access.

Re: The classifiers Anthropic puts in front of Fable are too zealous

#129
post #43

[flagged]

I havent used fable, but does it cost you the same when it downgrades to opus?

With the subscription, it costs less to use opus in that it doesn't chew up our session however the cost/benefit is balanced against not performing as well on certain tasks. So it's not a straight up yes/no.

Re: The classifiers Anthropic puts in front of Fable are too zealous

#130
post #90
post #46

Earlier quoted context omitted.

It feels like the longtermist believers got involved in this (those are the people obsessed with garage-engineered designer viruses who have a very tenuous grasp on how biology research actually works).

I assumed they just wanted to cultivate FOMO to sell an even more expensive version to researchers later on.

Don't they already do that?
Post reply on HN