Earlier quoted context omitted.
China aren't offering a cheaper solution. They are subsidizing an existing one (which is already subsidized) in order to gain foothold. The difference is that in the US subsidies come from VC, while OP implies subsidies come from the AI labs that buy the training data (which may as well also be VC backed, so just one extra hop). This isn't "the market working as intended", this is an exhaustion fight to the bottom wh…
> China aren't offering a cheaper solution. They are subsidizing an existing one (which is already subsidized) in order to gain foothold. In my economics classes, we were told that (in a "free market" argument) the best thing to do if a subsidy is making something you want cheaper is to use it. You're getting your thing, and at a reduced cost. (I'm not really replying to you per se, I'm curious how "free market" folk…
Anthropic says Alibaba illicitly extracted Claude AI model capabilities
471–480 of 1001 posts
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#472There's two basic kinds of distillation: 1) the massive [and dumb] method where you ask a question and use the answer as reinforcement (Black Box), and 2) more targeted distillation where you use one model to directly inform/train/guide another model (RLAIF). The latter is basically fine-tuning the model with direction from another model. Thousands of businesses do this every day to fine-tune. This is almost certainl…
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#473Here's what is happening: Chinese resellers are offering Claude tokens at 70-90% below official Anthropic API prices. They achieve this by reselling capacity from pooled Claude Max accounts, payments fraud, and also reselling the model output & reasoning chains to various Chinese labs. They are subsidizing model access in exchange for user logs and reasoning traces, which they then sell as training data, allowing the…
Don’t put that on Chinese.
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#474Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#475Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#476> The strike by Alibaba is described as a "distillation" effort, which Anthropic has said involves training a less capable model on the outputs of a stronger one. I don't see what's wrong about this. > Anthropic said the campaign was conducted between April 22 and June 5, 2026, and generated more than 28.8 million exchanges with Claude through almost 25,000 fraudulent accounts. What makes the accounts fraudulent? If…
Because Anthropic has terms of service with more stipulations than just "you must pay and can use the service for any purpose"?
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#477Back in the day, an "attack" was supposed to mean be someone acquiring our assets without paying for them or without having our consent. But none of this seems to have happened in this case.
We built a product without paying for most of the raw material we have used, and we don't call that as an "attack". Did we change the meaning of "attack"?
Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#478Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#479Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities
#480There's two basic kinds of distillation: 1) the massive [and dumb] method where you ask a question and use the answer as reinforcement (Black Box), and 2) more targeted distillation where you use one model to directly inform/train/guide another model (RLAIF). The latter is basically fine-tuning the model with direction from another model. Thousands of businesses do this every day to fine-tune. This is almost certainl…
This is, in part, a problem every judicial and legislative system has faced since forever: form versus function.
Take a classic elicitation spying techniques: a foreign spy meets a military officer/scientist at a bar, strikes up a conversation, makes an observation wondering how could a missile hit some target at some accuracy and elicits a response that with laser guidance it is entirely possible. From this they get info that there is some technology to laser guide missiles. Or in retail, a competitor hiring a secret buyer for core baskets of goods and analyzing prices in the receipts.
The function is espionage, the form is conversation and all info is in a sense provided willingly. Where do you pull the slider?
These distillation "attacks" are not only indistinguishable from evals, they ARE evals. The function is own model training, the form is eval. Normally, one would expect to have risk benefit analysis based discussion which direction to push the legality slider to. The problem with these recurring statements is that they invoke enshitification of legislature.