Live data from Hacker News

The Kimi K3 Moment

stephen.bochinski.dev

641–644 of 644 posts

Re: The Kimi K3 Moment

#641

Earlier quoted context omitted.

I am not arguing that Kimi did not distill Opus or Sonnet (which they almost certainly have, as opposed to Fable or Sol). I am arguing that even if they did not distill, the model could still identify itself as Opus or Sonnet. That being said, I'd like to point out a few things: - The second link (to claude-sonnet-4-2025051) had 17k matches, not just 3k. - grep.app does not index the entirety of GitHub, so the real n…

Your theory doesn't hold up against the data. While model identifiers like "claude-opus-4-5-20251101" appear thousands of times across GitHub and other code sources, so do "gpt-4o-2024-08-06", "gpt-4.1-mini-2025-04-14", and "gemini-1.5-pro-002" in similar amounts. I clicked your links and examined the GitHub configuration files. There are hundreds of other model names that all appear in these config files at comparab…

I would not put any trust into that AI-generated "research". It does not control for how OpenAI, Anthropic and Moonshot are doing tokenization and token healing, so the results are meaningless.

Re: The Kimi K3 Moment

#642

Regardless of whether they achieved parity via distillation, or whether they got here via independently constructing a model from scratch, it was always going to end this way for the frontier American labs. Distillation “attacks” are not attacks. The frontier labs “distilled” all existing human written knowledge into their models, there was always going to be a second class lab that would distill that model into a ch…

I saw your post about the Kimi K3 Moment and how you mentioned the pricing of models from Anthropic, OpenAI, and Kimi, I thought you might be interested in ShipDiff, which could help with tracking competitor pricing strategies, feel free to ignore this if not relevant

Re: The Kimi K3 Moment

#643
post #613

Earlier quoted context omitted.

> Service Misuse. You acknowledge that without the written consent of us and/or the relevant rights holders, (i)you have no authority to use Kimi and the content generated by Kimi in any commercial manner; (ii)you may not use our Services to develop products or services that compete with us. https://www.kimi.com/user/agreement/modelUse

That’s for the consumer app / chatbot. For the api the terms are different: https://platform.kimi.ai/docs/agreement/modeluse

Good to know thanks for the clarification

Re: The Kimi K3 Moment

#644

Earlier quoted context omitted.

my first prompt to any Kimi model was K3 via Pi, some version of "hi kimi!!" and the response was telling me "I'm actually Claude." this is not hard to repro, just use a system prompt that doesn't mention the model name. that said, if they bootstrapped with opus 4.6 convo sft data they had sitting around... so what?

The main story is what isn't being talked about. Chinese labs exfiltrated trillions of tokens of high-quality output from Anthropic and OpenAI, through proxies and heavily discounted token resellers, which they distilled and used for training data for their own models. Instead of spending 12-18 months building their own robust harnesses and painstakingly creating quality training data (which is what Anthropic and Ope…

Not at all curious you keep citing sources only from a platform run by an incompetent rascist arse wipe that you constantly boost all over here.

"The community is happy to overlook any questionable methods Chinese labs used to build them."

Oh and you aren't willing to overlook all the actual illegal activities, as in actual court cases that the US labs and corps used to train their models?

Industrial sabotage? Pa-lease! get off your pretend high horse and stop making us laugh.

Post reply on HN