There's two basic kinds of distillation: 1) the massive [and dumb] method where you ask a question and use the answer as reinforcement (Black Box), and 2) more targeted distillation where you use one model to directly inform/train/guide another model (RLAIF). The latter is basically fine-tuning the model with direction from another model. Thousands of businesses do this every day to fine-tune. This is almost certainl…
>But if you show them a jailbreak of their model that bypasses their safety, they'll tell you that any model can eventually be jailbroken so don't worry about safety. Yes this is in line with what Anthropic said in their public statements about their Fable access restriction by the government directive. The hypocrisy and inconsistency in their statements and behavior feels quite childish and controlling. I believe ou…
Anthropic says Alibaba illicitly extracted Claude AI model capabilities
391–400 of 1001 posts
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#392Earlier quoted context omitted.
China aren't offering a cheaper solution. They are subsidizing an existing one (which is already subsidized) in order to gain foothold. The difference is that in the US subsidies come from VC, while OP implies subsidies come from the AI labs that buy the training data (which may as well also be VC backed, so just one extra hop). This isn't "the market working as intended", this is an exhaustion fight to the bottom wh…
I mean, for what it's worth, we have subsidized Anthropic by allowing them to train on copyrighted stuff. (I know it is still legal, and I support the legality, but the economics are what they are with people's content paying a big one time subsidized cost (to the level of at least 500B). So, the least Anthropic can do is pay it forward.
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#393Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#394>The strike by Alibaba is described as a "distillation" effort, which Anthropic has said involves training a less capable model on the outputs of a stronger one. Claude used TB of content without permission to train their model and it was ok for them. Now someone else uses the output of a Claude model to train model and they cry foul.
It was not okay for them, they had to pay one billion dollars.
Essentially peanuts compared to what they would have to pay to obtain the rights of everything they pirated.
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#395Earlier quoted context omitted.
>This is great for competition! Chinese vendors offering a cheaper solution = what economics told me the free market was all about. Yeah, like all those Chinese bootleggers selling DVDs for a few dollars rather than $20. Free market! https://news.ycombinator.com/item?id=48664814
The output of Claude is not eligible for copyright protection. I'm not sure how the analogy of bootlegging DVDs would work, given that.
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#396Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#397I wouldn’t shout too loud if I were the second kid.
They’ve shouted before about the dangers of their favorite toy and the teacher took it away from them, lets see what happens if they shout too much this time.
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#398Here's what is happening: Chinese resellers are offering Claude tokens at 70-90% below official Anthropic API prices. They achieve this by reselling capacity from pooled Claude Max accounts, payments fraud, and also reselling the model output & reasoning chains to various Chinese labs. They are subsidizing model access in exchange for user logs and reasoning traces, which they then sell as training data, allowing the…
How dare they. Only Anthropic is allowed to sell its tokens at 70-90% below the API prices.
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#399Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#400Here's what is happening: Chinese resellers are offering Claude tokens at 70-90% below official Anthropic API prices. They achieve this by reselling capacity from pooled Claude Max accounts, payments fraud, and also reselling the model output & reasoning chains to various Chinese labs. They are subsidizing model access in exchange for user logs and reasoning traces, which they then sell as training data, allowing the…
Those resellers are simply just selling Kimi K2.5 or GLM5.1 as counterfeit Opus. We, Chinese, know how to play the counterfeit game for a long time in so many industry.