Live data from Hacker News

Anthropic says Alibaba illicitly extracted Claude AI model capabilities

reuters.com

881–890 of 1001 posts

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#881
post #674

The hypocrisy of Anthropic complaining about "illicitly extracting its Claude AI model capabilities" and supporting the White House's accusation of China "stealing U.S. AI labs' intellectual property on an industrial scale" is hilarious. Anthropic, OpenAI, Google, Microsoft, et al trained their models by ignoring the rights of copyright holders when harvesting whatever content they could. Now one of them is crying fo…

Not really. Data mining for AI is presumably fair use, whereas when you sign up for a Claude account, you enter into a legally binding contract that says you will not distill a model based on its outputs.

“Legally binding” bs that a judge would laugh off.

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#883
If the concern is that China is catching up on model capabilities (which is only a big deal if you lean in to adversarial geopolitical zero-sum thinking), the fact that they're using American models to train theirs should give people comfort that they're nowhere near the cutting edge

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#884

Earlier quoted context omitted.

> China aren't offering a cheaper solution. They are subsidizing an existing one Chinese labs are also pursuing legit frontier-advancing R&D into efficiency and publishing papers in the open, a culture that's in retreat at top American AI labs

Their is plenty of innovation happening on both sides of the Pacific. Again, China publishes open source because they don't have another game they can play. They distill because they don't have the compute to compete. They are great lab, for sure, but the fundamentals are driving their behavior.

The fact that are people that genuinely believe you can train an LLM by using random QAs obtained from another LLM is astonishing. Let alone the fact that it makes absolutely zero financial sense.

At this point this is being repeated so often that completely uninformed users are taking this at face value.

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#885

Earlier quoted context omitted.

Agreed! I had to do a double take and check the URL. I thought I am reading a press release rather than actual reporting.

https://news.ycombinator.com/item?id=13155538

ironically, I think this is why the jobs apocolypse is overblown, Ai is only good at a thing if the people using it are also good at that thing, and people are attributing Ai as superhuman at things they do not know themselves

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#887
post #807

“Distillation attack” are we joking here. If anything these models should be compelled to be public since they have been trained off public data. What an absurd overreach to call this an attack. It’s clear they are scapegoating national security and China at this point to build an anti-competitive moat. I generally really like Anthropic’s work and models but stuff like this scares me for the future. We are positionin…

The core of the training data is public, but the part that actually makes these models smart came from (pretty highly-paid) experts via platforms like Mercor. Claude didn't magically learn to write good code by reading all of GitHub - humans trained it in that, more or less manually.

Given the breadth of LLM knowledge, I somehow doubt this. Sure, it’s probably responsible for the quality of LLM insights, but I don’t think anyone was asking experts about e.g. the complex ecological effects of invasive zebra mussels and their provenance in Lake Michigan.

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#888
post #359

Earlier quoted context omitted.

That's a great cost-benefit ratio. Can you and I steal and do illegal things and pay the same cost?

Sure, but only if you get the same benefits

looks like we can't today. Man it would be great to figure out how to be above the law just like how these other rich people in different social classes are.

Re: Anthropic says Alibaba illicitly extracted Claude AI model capabilities

#889

Earlier quoted context omitted.

> If anything these models should be compelled to be public since they have been trained off public data I'm starting to come around to this idea TBH. For a while my position was: "these companies have invested billions into training these models, therefore they should be able to control them and profit off them" but looking deeper at where they got their training data, my view is starting to shift. IMHO I feel like…

labs invest multiple billion dollars a year each in private data, and that number is growing. internet training data is not where frontier capabilities come from, this view is outdated

Why are the leading models capable of regurgitating full copyrighted works such as "Harry Potter" and "On the Road"? Did they hire someone to type those out for them?

https://arxiv.org/abs/2601.02671

Post reply on HN