Writing fingerprint analysis of responses reveals Kimi's similarity to Claude
1–10 of 73 posts
Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude
#2There is an optimal answer to any question. Something that maximizes utility and minimizes tokens.
Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude
#3Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude
#4Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude
#5Stolen data was stolen. Oh no! Anyway.
This looks very similar to the claim that distilling a model from Anthropic is the same thing as Anthropic distilling the model from information on the internet.
Which is very flawed, since distillation requires the thing to exist in the thing it’s distilled from. And no LLM model existed in the information Anthropic used to train the model. Instead the model was built using information and utilizing new technology including hardware, software, transformer architecture, etc.
Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude
#6How do you do cross entropy analysis for Claude without logits?
Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude
#7Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude
#8How do you do cross entropy analysis for Claude without logits?
Edit: No that doesn't seem to be what's happening here. I think it's some kind of word frequency analysis.
Re: Writing fingerprint analysis of responses reveals Kimi's similarity to Claude
#9But I don’t grok training enough to know that’s silly.
If my new prior is you can…that’s a pretty thin moat that’s essentially indefensible.