Live data from Hacker News

Anthropic says Alibaba illicitly extracted Claude AI model capabilities

reuters.com

321–330 of 1001 posts

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#321
post #282

Earlier quoted context omitted.

[flagged]

The issues with LLMs go beyond just IP theft. I would not say PRC making LLMs cheaper is the best outcome (though it is better than nothing). The best outcome would be to make the practice of training on our data without consent illegal, which would simultaneously slow down economic change and make it more organic as well as give PRC companies less capabilities to extract.

> The issues with LLMs go beyond just IP theft.

There is no IP theft because LLM outputs aren't protected, just egregious ToS violations.

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#322

There's two basic kinds of distillation: 1) the massive [and dumb] method where you ask a question and use the answer as reinforcement (Black Box), and 2) more targeted distillation where you use one model to directly inform/train/guide another model (RLAIF). The latter is basically fine-tuning the model with direction from another model. Thousands of businesses do this every day to fine-tune. This is almost certainl…

https://research.nvidia.com/labs/lpr/slm-agents/ - Distillation data is a natural byproduct of using these models. There's no effective defence against it. Anthropic is degrading thinking blocks to summaries to slow it down and hide model internals, but in the end, the math says you're SOL and it works on MNC/Large Corporate scale well enough that the moment cost becomes a priority, you're left without the lock in you need to keep customers paying.

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#323

Here's what is happening: Chinese resellers are offering Claude tokens at 70-90% below official Anthropic API prices. They achieve this by reselling capacity from pooled Claude Max accounts, payments fraud, and also reselling the model output & reasoning chains to various Chinese labs. They are subsidizing model access in exchange for user logs and reasoning traces, which they then sell as training data, allowing the…

> They achieve this by reselling capacity from pooled Claude Max accounts, payments fraud, and also reselling the model output & reasoning chains to various Chinese labs.

Claude never provides the raw reasoning chain. What you see is just a summary of that reasoning. Getting the full thinking output requires an enterprise agreement.

https://patrickmccanna.net/the-text-in-claude-codes-extended...

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#325

Here's what is happening: Chinese resellers are offering Claude tokens at 70-90% below official Anthropic API prices. They achieve this by reselling capacity from pooled Claude Max accounts, payments fraud, and also reselling the model output & reasoning chains to various Chinese labs. They are subsidizing model access in exchange for user logs and reasoning traces, which they then sell as training data, allowing the…

> which is also why Anthropic introduced identity verification, to slow down the onslaught of bots. Lol. The irony is thick for anyone who ever had to attempt defense against an onslaught of American AI lab crawlers that ignore robots.txt

Yeah nobody is gonna be shedding any tears for them

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#326
post #282

Here's what is happening: Chinese resellers are offering Claude tokens at 70-90% below official Anthropic API prices. They achieve this by reselling capacity from pooled Claude Max accounts, payments fraud, and also reselling the model output & reasoning chains to various Chinese labs. They are subsidizing model access in exchange for user logs and reasoning traces, which they then sell as training data, allowing the…

[flagged]

I hope you say the same when these cheap llms are used in drones to target humans. The world models are exactly built with that direction in mind.

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#327

The hypocrisy of Anthropic complaining about "illicitly extracting its Claude AI model capabilities" and supporting the White House's accusation of China "stealing U.S. AI labs' intellectual property on an industrial scale" is hilarious. Anthropic, OpenAI, Google, Microsoft, et al trained their models by ignoring the rights of copyright holders when harvesting whatever content they could. Now one of them is crying fo…

It's not exactly the same, since any Claude output is public domain under current law. So the Chinese aren't stealing anything here.

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#329
post #326
post #282

Earlier quoted context omitted.

[flagged]

I hope you say the same when these cheap llms are used in drones to target humans. The world models are exactly built with that direction in mind.

Cool beans boomer alarmist stance. The Chinese models here are doing what they’re supposed to price the market accordingly.

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#330
post #37

Earlier quoted context omitted.

Why do you need macbooks? Just rent servers from any hosting provider.

Not going to work for very long or at any scale coming from datacenter/hosting provider IPs. Google "residential proxies for sale" for the tip of an iceberg of how they snowshoe the traffic.

Hey, if the bastards can use residential IPs to suck all information into their models with their crawlers, so can we!
Post reply on HN