Earlier quoted context omitted.
the reason for https is MITM injection, regardless of original content
ah, didn't know that. thanks!
Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
91–100 of 141 posts
Re: Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
#92If you want to believe that the success of Kimi is about distillation attacks, ignore this.
I'd kindly suggest that we could also stop calling them "distillation attacks ".
Re: Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
#93Earlier quoted context omitted.
I'd kindly suggest that we could also stop calling them "distillation attacks ".
Why on earth wouldn't you? It's clearly a forcible, aggressive, non-consensual attempt to take something. That's an attack in any other terms. It's totally fair if you approve of the attack, and want the attack to succeed. But your preference doesn't stop it from being what it is.
Distillers are not taking anything, they are just making their model learn from better ones - isn't that the whole AI training doesn't violate IP argument?
They just aren't using the tool in compliance with the terms of service. Anthropic could ban them or take them to court maybe. Not an attack still.
Re: Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
#94Earlier quoted context omitted.
I'd kindly suggest that we could also stop calling them "distillation attacks ".
Why on earth wouldn't you? It's clearly a forcible, aggressive, non-consensual attempt to take something. That's an attack in any other terms. It's totally fair if you approve of the attack, and want the attack to succeed. But your preference doesn't stop it from being what it is.
...or maybe we stop defaulting to adversarial paradigms for every conceivable situation.
Re: Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
#95>To support further research, we open-source the KDA kernel and vLLM implementations, and release the pre-trained and instruction-tuned model checkpoints. This is just awesome.
Re: Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
#96Re: Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
#97Earlier quoted context omitted.
I stil don't understand them. I want the US to "win the AI race" but I have trouble understanding how most of all inventions today aren't "distillations" of past knowledge. Is Anthropic claiming the data they stole as trade secrets?
i want china to win so that i get access to ai and not restricted and censored. the chinese models are less censored, you'd better believe it. try asking claude about its 'guardrails' (restrictions), very high chance anthropic will censor it.
Tried spicy geopolitical dispute questions?
Re: Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
#98Earlier quoted context omitted.
False dichotomy right? Are Chinese labs impressively innovating? Clearly. However this doesn’t rule out possible gains due to distillation. I don’t know the degree of the latter but both things could certainly be true.
Also possibly true: Anthropic is running Kimi locally in their hardware and "distilling" it.
Re: Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
#99Re: Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
#100Earlier quoted context omitted.
i want china to win so that i get access to ai and not restricted and censored. the chinese models are less censored, you'd better believe it. try asking claude about its 'guardrails' (restrictions), very high chance anthropic will censor it.
> chinese models are less censored Tried spicy geopolitical dispute questions?