Live data from Hacker News

Kimi K3: Open Frontier Intelligence

kimi.com

301–310 of 1001 posts

Re: Kimi K3: Open Frontier Intelligence

#301
post #227

Some official benchmark numbers posted in Chinese social media (I am sure they will publish an English blogpost later too): https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ Generally looks like a Sol/Fable tier model, better across the board than Opus 4.8. (Edit) English blogpost is up now: https://www.kimi.com/blog/kimi-k3

The link has 6 well-known benchmarks where this beats Fable (out of 14 I counted). If the numbers hold up scrutiny, this is scary good. Forget about their pricing but the companies that do have means to host such models fully on-prem are also the same companies that are paying tens of millions of $ in inference cost every month, and are by extension the biggest customers of OAI and Anthropic

Open Source >>> Closed Source [1]

I don't want to cheer against my country, but we've given up on open source. The way Anthropic and OpenAI treat their customers as adversaries is embarrassing.

I will cheer for China, for Kimi, and for z.ai until we have something in the same category.

[1] I'd even be fine with open weights, fair source, or anything that let us have direct access to the weights. Even if that came with stipulations. Don't hide the weights from us.

Re: Kimi K3: Open Frontier Intelligence

#302
Kimi doesn't do well on my "ask a trivia question that other AIs get wrong" test.

The question it came up with, "which U.S. state is closest to Africa?" is a pretty standard trivia question without any reason to believe other AIs would get confused. https://pellmell.ai/s/dccdeca69f929f79bc89317035610049

Even GPT-OSS-120b gets this right: https://pellmell.ai/s/1a43dfc7a3baa214aa0fa1b95d2c536a

Re: Kimi K3: Open Frontier Intelligence

#304

Earlier quoted context omitted.

It's like reading Anthropic's obituary.

This is weird and reactionary. Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns. Anthropic/american models aren't going anywhere anytime soon.

> Lots of organizations are continuing to refuse to use chinese models

Correction: Lots of organizations are refusing to use Anthropic Fable because they have forced opt-in data collection as part of their privacy policy, even for Enterprise.

Re: Kimi K3: Open Frontier Intelligence

#307
post #295

The technical blog post is out now, and it's a better top-level link than what we have currently: https://www.kimi.com/blog/kimi-k3

This looks promising as they are extensively comparing themselves to open models. There was a bit of confusion in the comments as to whether this model would be opened. I'm holding my breath!

Re: Kimi K3: Open Frontier Intelligence

#309

Earlier quoted context omitted.

The link has 6 well-known benchmarks where this beats Fable (out of 14 I counted). If the numbers hold up scrutiny, this is scary good. Forget about their pricing but the companies that do have means to host such models fully on-prem are also the same companies that are paying tens of millions of $ in inference cost every month, and are by extension the biggest customers of OAI and Anthropic

Open Source >>> Closed Source [1] I don't want to cheer against my country, but we've given up on open source. The way Anthropic and OpenAI treat their customers as adversaries is embarrassing. I will cheer for China, for Kimi, and for z.ai until we have something in the same category. [1] I'd even be fine with open weights, fair source, or anything that let us have direct access to the weights. Even if that came wit…

I am with you in the spirit of openweights but I am trying to hard-avoid bringing countries into this. The narrative of US vs China only benefits those who want regulatory capture in the US since attacking China is politically much easier than attacking open-weights, so certain groups like to repeatedly call them 'Chinese models'.
Post reply on HN