Live data from Hacker News

Kimi K3: Open Frontier Intelligence

kimi.com

631–640 of 1001 posts

Re: Kimi K3: Open Frontier Intelligence

#631
post #37

> In our evaluations, Kimi K3 delivers frontier-level performance. Among the models tested, its overall intelligence ranks second only to Claude Fable 5 and GPT-5.6 Sol. For the complete benchmark results, see our tech blog. The full model weights of Kimi K3 will be released in the coming days. More details on the architecture, training, and evaluation will be published together with the Kimi K3 technical report. > K…

> Maybe another DeepSeek moment right here. Surely not... What made DeepSeek disruptive was that the cost was 10X lower. In this case, the cost is about 2X lower the Sol I think? At 2X, you're pretty close to the error margins due to token efficiency etc... I'd say this is "on trend" for open models catching up to frontier labs, but its not a "change in the trend" like DeepSeek was IMO.

If AA is to be believed then per-task it is about the same cost as Sol. Agree that it's very different from DeepSeek v4 Pro, which is ~15x cheaper than K3.

https://artificialanalysis.ai/models/kimi-k3#price-cost

Re: Kimi K3: Open Frontier Intelligence

#632

So Chinese labs are driving essentially towards commodotized intelligence. Even if its a few months behind the US. Is this a classic 'commoditize my compliment' situation? They want to sell the hardware and infrastructure behind AI and make the software part not the value driver / moat? I can see it. But also even two Chinese labs sinking 100s of millions USD into training isn't exactly commoditization. It's still a…

National strategy. Thanks to smart investments, China's energy costs are far lower than everybody else's. If intelligence becomes a commodity, then China will be providing it for the world. Incredible leverage.

Re: Kimi K3: Open Frontier Intelligence

#633

Earlier quoted context omitted.

> Maybe another DeepSeek moment right here. Surely not... What made DeepSeek disruptive was that the cost was 10X lower. In this case, the cost is about 2X lower the Sol I think? At 2X, you're pretty close to the error margins due to token efficiency etc... I'd say this is "on trend" for open models catching up to frontier labs, but its not a "change in the trend" like DeepSeek was IMO.

cost has nothing to do with why deepseek was disruptive, the fact that it means there is zero moat around anthropic or openai is what's disruptive about it. it means in the mid-term LLMs will be commoditized and customers will flock to the cheapest inference wherever they can find it. there's no reason to stick to the "frontier" labs

> cost has nothing to do with it

> customers will flock to the cheapest inference

Re: Kimi K3: Open Frontier Intelligence

#636
post #177

Earlier quoted context omitted.

I wouldn't be surprised if models were optimizing for rendering SVG pelicans at this point

every ai release thread seems to have this same sequence of comments

what if the real pelican was that chain of comments

Re: Kimi K3: Open Frontier Intelligence

#637
post #518

Earlier quoted context omitted.

If there was some grand strategy for all Chinese labs, surely it'd have leaked by now. I think its more likely that: - Companies can still make money from commodities - Chinese labs only have 5-10% the valuation of OpenAI/Anthropic, so massive monopoly profits aren't necessary. Profit expectations for tech companies in China are really low in general, complete opposite of the US. - Open weighting is a great way to ge…

- AI investment is basically the only thing keeping the American economy treading water at the moment, and kicking the chair out from under that industry benefits China tremendously. I'm definitely not saying that's the only factor, but I think it's naive to assume it isn't at all a factor.

It's definitely a factor.

In the same way that the US going to war with an important Chinese oil supplier: 85% of Iranian oil going to China and comprising ~14% of China's oil imports.

There's always geopolitical reason behind the reason.

Re: Kimi K3: Open Frontier Intelligence

#638

Earlier quoted context omitted.

It's the same reason Meta open sourced Llama and AMD open sourced FSR. When you're behind it is a prudent strategy because it undermines investment in the private frontier. Once you're on top you pull the rug and go closed source. There are no morals in this anywhere to be found. > 100s of Millions That is utter peanuts given the stakes. This is competition between two super powers for the most important technology i…

> This is competition between two super powers for the most important technology in human history. Photonic computing?

I'd vote fire, but people say Grug brain ethr1 know too much about old technology.

Re: Kimi K3: Open Frontier Intelligence

#639

Earlier quoted context omitted.

The link has 6 well-known benchmarks where this beats Fable (out of 14 I counted). If the numbers hold up scrutiny, this is scary good. Forget about their pricing but the companies that do have means to host such models fully on-prem are also the same companies that are paying tens of millions of $ in inference cost every month, and are by extension the biggest customers of OAI and Anthropic

Open Source >>> Closed Source [1] I don't want to cheer against my country, but we've given up on open source. The way Anthropic and OpenAI treat their customers as adversaries is embarrassing. I will cheer for China, for Kimi, and for z.ai until we have something in the same category. [1] I'd even be fine with open weights, fair source, or anything that let us have direct access to the weights. Even if that came wit…

I just want to see dario cry for some reason . i cannot explain it but i want him in particular to lose.

Re: Kimi K3: Open Frontier Intelligence

#640
post #266

Earlier quoted context omitted.

GLM is actually quite expensive in actual practice because it's not very token efficient. I've yet to find a way to run it on a monthly sub reliably for cheaper than Codex. Neuralwatt was cheap (but slow) but they cranked their price. Ollama monthly sub is speedy but doesn't offer a lot of quota. Right now unless you're paying by the token, there's no cost based reason to use the open weight models for daily coding w…

Compared to the flagship models GLM is still a 1/10th the price on the task I have tested.

Only if you're paying for said flagship models through API prices. Which I specifically said I was not and specifically mentioned coding plans.
Post reply on HN