Live data from Hacker News

The Kimi K3 Moment

stephen.bochinski.dev

501–510 of 644 posts

Re: The Kimi K3 Moment

#501
post #487

Earlier quoted context omitted.

You skipped the part where Ford buys the cars at 90% off and sells them at 80% off, at a profit. Then gets paid by competitors for the driving data. At the same time, Volvo is running the exact same hustle, except they buy the cars with stolen credit cards, so they get the cars for free.

Don’t sell your cars at a loss then.

[deleted]

Re: The Kimi K3 Moment

#502

Earlier quoted context omitted.

I strongly agree with the premise that distillation is not an “attack”. But that said: K3 is not a distilled version of Fable or Sol. Fable has been barely available and Sol was just released! Moreover, K3 is superior to both models in some domains, according to user scoring on the Arena. API distillation can’t give you these results anyway. All it is useful for is bootstrapping RL in new domains to get past the “col…

API distillation doesn't have to explain all of K3's capabilities for it to have happened. Kimi K3 reproducibly identifies itself as Claude: https://x.com/denisewu/status/2077984660211269870 This behavior is exactly what you'd expect from a model distilled from Claude. There's a detailed analysis of K3's ambiguous identity here: https://github.com/rgreenblatt/which_claude_is_k3/blob/main/... This analysis observed K3…

Claude Sonnet 4.8 reproducibly identifies itself as DeepSeek when asked in Chinese:

https://x.com/stevibe/status/2026227392076018101

I mean, people can point fingers however they want, and the fact is nobody actually "owns" the data they feed to their LLMs...

Re: The Kimi K3 Moment

#503

Earlier quoted context omitted.

Sure, but then Qwen should leak that too, and it doesn't. K3 calls itself Claude 7 out of 48 times, Qwen does it 0 out of 48, and the only other model to identify itself as Claude is DeepSeek. and DeepSeek is alleged to also distill from Claude data anyway. So this isn't something every model absorbed from the same web text. And you skipped over the strongest datapoint that K3 is distilled: K3 reproduces Claude's pub…

Did you know claude models identify as qwen or deepseek when asked in chinese?

Supplementing with evidence: https://x.com/stevibe/status/2026227392076018101

Re: The Kimi K3 Moment

#504

This was always where this was heading, but we got here much faster than expected. Once western governments declare it to be a "national security" risk for citizens to have access to open-weight frontier models, and once they classify using these models as acts of terrorism, what will that world be like? Will using Kimi K3 come to be like how napster was in the olden days? Everybody knew it was technically illegal, b…

> Or it will be like cannabis, where a guy in the neighborhood will low key rent you metered access to the 8x5090 rig in his basement he cobbled together from parts on ebay?

https://www.youtube.com/shorts/iNotXHO8NWU

Re: The Kimi K3 Moment

#505
post #116

Earlier quoted context omitted.

At this point, the United States will lose that battle most Countries in the world are going end up using electronics from Asia, that ship has sailed Japan, China, Korea, Singapore, Taiwan, Vietnam, dominate that area, China already dominates EVs, Drones and many other electronic devices, and with the way Donald Trump has picked fights, Europe, Canada, Australia, New Zealand, Mexico and many others are looking for ot…

Singapore? Singapore doesn't produce shit.

Are you dissing my boy Sound Blaster??

Re: The Kimi K3 Moment

#506

Earlier quoted context omitted.

I've been in China in 2015 and like anywhere else in the world it was very mixed: some urban areas like central NY or central Madrid and Milan (or much shinier) and some rural areas like 200 year ago, but inevitably with electronics. Basically in every country of the world you can travel one hour from big cities and get in a place deep in the fields or the woods with very different needs and dynamics from the city. T…

I don’t know any cities in Americas that is comparable to Shanghai or even Japan in term of transportation or convenient

Paris has a metro station everywhere at least in what a tourist can assume to be an enlarged city center. Tokyo is another city with a lot of metro stations. Manhattan too, at least up to Central Park (but 20+ since my last visit.) I don't remember Shanghai to stand out positively or negatively, but 11 years can be a long time.

Re: The Kimi K3 Moment

#507

Even in this very thread the feedback on Kimi's actual efficacy is debated. I personally feel its worse than both Fable and 5.6 Sol, but I feel like the conversation isn't really about whether its good or not, but a backlash against the U.S governments foray into regulation. So I think people _want_ it to be superior out of anger/frustration with the current situation.

This seems like a replay of what happened with DeepSeek. They put out v3, or whichever one it was, and everyone said it was over for US companies... then everything continued on.

The markets can be irrational for a while, but if the Chinese models are about 90% of the performance of OpenAI and Anthropic models, and the Chinese companies are This isn't just the AI race, but the end to perceived American exceptionalism (where USA wins by default). It's going to take a while for people to recognize that. Before that the markets will still go crazy, but that's not evidence things will continue on as "normal".

Re: The Kimi K3 Moment

#508

Regardless of whether they achieved parity via distillation, or whether they got here via independently constructing a model from scratch, it was always going to end this way for the frontier American labs. Distillation “attacks” are not attacks. The frontier labs “distilled” all existing human written knowledge into their models, there was always going to be a second class lab that would distill that model into a ch…

if distillation is the key, why the fuck all other competitors do not release competitive models? and only Chinese can distill this great?! Am I smoking too much?

Re: The Kimi K3 Moment

#509
post #73

Well, there is the small issue of privacy policy: Kimi will train their models on your interactions if you use their subscriptions, and only with direct API usage (billed at API prices) they say they won't. Whether you trust that is another matter. Those things do make a difference to some of us, even though nothing is black and white. In my case, I'll probably want to wait until other providers appear through OpenRo…

> Kimi will train their models on your interactions I find these kinds of concerns increasingly silly: most of the input to these models will be ... previous output from the very same models, alongside the occasional half-assed human command to fix something and "make zero mistakes". Who cares if they train on that? Let them, if it makes their future models better! 99% of users are not working on any special IP to wo…

It means there's a non-trivial chance a future version of the model will know private information about you.

Maybe you're super careful with this stuff, but with agents and harnesses being given access to user data and accounts, I don't think it's feasible to actually monitor what information is uploaded and whether they involve private information.

I personally keep local models around because of this.

Post reply on HN