Live data from Hacker News

Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

typebulb.com

51–60 of 73 posts

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#51

This data sort of disqualifies itself: unless Moonshot has a time machine, K3 should be more similar to Opus 4.5-4.8 than Fable 5. Keep in mind, Anthropic started limiting access and introduced anti-distillation measures around 4.5-4.6 (?). So the majority of distillation should have happened on earlier models. Maybe a better explanation is that they have access to the same training datasets? Which if private can aga…

[deleted]

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#52
post #14

Earlier quoted context omitted.

What of the costs for creating the data that was used to train the model being distilled?

Do you people seriously believe that Moonshot and the Chinese personal-cult-state dont possess and train on the same torrents?

They aren’t hypocritically crying foul about it, so no we don’t care if they do.

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#53
post #35

Earlier quoted context omitted.

Does distilling actually violate any IP protection laws? Sure, it’s against their terms of service.

I get what you're saying and two wrongs do not make a right but the irony, and why people are even talking about this, is that the thing being distilled clearly, and knowingly, violated copyright & terms across the entire internet.

Yeah, everyone is saying that but I don't find the irony very interesting.

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#54
post #10

All frontier models have been trained without any regards for IP protection laws. I don't see how anyone can argue in good faith that distillation is not fair game and does not ultimately "benefit humanity™"

For me, I don't really care about the theft aspect but when people are claiming that these open models are better value or going to overtake anthropic/openai models, the implication that the open models are training of distilled data means all the "progress" they are making is just mimiced from the closed models. It's a bit interesting how the open models are able to keep pace with the closed models except whole main…

Distillation is absolutely not the reason they're good. It's not necessarily even done on a more capable model. It can even be done on itself and still bring improvement, or on a weaker model as well (see GLM and Gemini, which is definitely true because it repeats Deepmind's injections).

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#55
post #10

All frontier models have been trained without any regards for IP protection laws. I don't see how anyone can argue in good faith that distillation is not fair game and does not ultimately "benefit humanity™"

Does distilling actually violate any IP protection laws? Sure, it’s against their terms of service.

If it does then their original training did as well.

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#56
post #52

Earlier quoted context omitted.

Do you people seriously believe that Moonshot and the Chinese personal-cult-state dont possess and train on the same torrents?

They aren’t hypocritically crying foul about it, so no we don’t care if they do.

[deleted]

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#57
post #10

All frontier models have been trained without any regards for IP protection laws. I don't see how anyone can argue in good faith that distillation is not fair game and does not ultimately "benefit humanity™"

For me, I don't really care about the theft aspect but when people are claiming that these open models are better value or going to overtake anthropic/openai models, the implication that the open models are training of distilled data means all the "progress" they are making is just mimiced from the closed models. It's a bit interesting how the open models are able to keep pace with the closed models except whole main…

It's important to note that even if these open models are distilled, they are showing genuine improvements in their architecture, which enables inference costs to be several factors below what equivalent closed models have.

The interesting question is: will Anthropic release a Fable like model with an architecture similar to Kimi, and get the inference cost gains? They should surely beat Kimi because they can internally distill as much as they want.

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#58
post #10

All frontier models have been trained without any regards for IP protection laws. I don't see how anyone can argue in good faith that distillation is not fair game and does not ultimately "benefit humanity™"

Training a model is very expensive and creates something no individual rights-holder could. Distilling a model copies this value add and captures it without bearing the cost that created it.

There is a lot more one has to be good at to make a fable level model. Distillation won't get you there, it will help refine some at the end

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#59
post #18
post #14

Earlier quoted context omitted.

What of the costs for creating the data that was used to train the model being distilled?

At least 1.5B by stealing books, per a recent ruling. I don’t feel sorry for the model companies

This actually undermines the argument that distilling is harmless because its founded on the idea that Anthropic did the same thing and didn’t have any repercussions.

Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude

#60
post #14

Earlier quoted context omitted.

What of the costs for creating the data that was used to train the model being distilled?

Do you people seriously believe that Moonshot and the Chinese personal-cult-state dont possess and train on the same torrents?

Well of course. But they dont hate Chinese model companies.
Post reply on HN