Live data from Hacker News

Alleged Distillation Attacks by DeepSeek, Moonshot AI, and MiniMax

twitter.com

11–17 of 17 posts

Re: Alleged Distillation Attacks by DeepSeek, Moonshot AI, and MiniMax

#12

Pot, meet kettle! I don’t think I’m the only one feeling some schadenfreude at this news. I suppose it’s ok when you’re a hot Silicon Valley scale-up to slurp up the rest of the worlds data for free and then hire hot shot lawyers to defend you against all the creatives you ripped off, but when it’s the “evil” Chinese doing the same to you it’s a dastardly “attack”?

Yeah - not only have we seend some of the same large companies that have trampled regular people and made examples of them in name of defending copyright fully ignore it when it was time to feed their AI models.

And now the hypocrisy went full circle with complains of others not respecting their rights!

Re: Alleged Distillation Attacks by DeepSeek, Moonshot AI, and MiniMax

#13
One difference between Anthropic and others is that Anthropic is crawling publicly visible information, and their argument is that this is fair use. Whereas these Chinese LLMs are circumventing an account creating process and terms of service to misuse non public information.

Lots of people think Anthropic training their own LLM is the same but it really isn’t.

Re: Alleged Distillation Attacks by DeepSeek, Moonshot AI, and MiniMax

#16
post #5
post #3

Earlier quoted context omitted.

From the tweet, Anthropic's point is that distillation is Ok, unless new model has safeguards removed or used for military or surveillance purposes.

The fact that they're calling it an "attack" implies otherwise. I find the entire premise of this announcement absurd. Fraudulent accounts? They're just accounts. They paid for the access the same as any other. They're accessing Claude just like a human (or *claw) would. There's no argument against their strategy that doesn't make them complete hypocrites in respect to how they got the model training data in the firs…

I agree with you, especially with this:

They paid for the access the same as any other.

If anything, this makes them more legit than Anthropic because they are paying for the content, whereas Anthropic just stole *all* the data they got a hold of. So, in this case the Chinese AI labs stand on higher moral ground LOL.

Re: Alleged Distillation Attacks by DeepSeek, Moonshot AI, and MiniMax

#17
If just 16 million examples were enough to significantly boost model quality (as Anthropic claims), it turns out that data quality beats quantity

Instead of vacuuming petabytes of trash from Common Crawl, you can just take high-quality distillate from a SOTA model and get comparable results. Bad news for anyone betting solely on massive compute clusters and closed datasets

Post reply on HN