Live data from Hacker News

Kimi K3: Open Frontier Intelligence

kimi.com

361–370 of 1001 posts

Re: Kimi K3: Open Frontier Intelligence

#361
post #227

Some official benchmark numbers posted in Chinese social media (I am sure they will publish an English blogpost later too): https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ Generally looks like a Sol/Fable tier model, better across the board than Opus 4.8. (Edit) English blogpost is up now: https://www.kimi.com/blog/kimi-k3

It's like reading Anthropic's obituary.

Nah:

https://www.youtube.com/watch?v=LSlV206xPqM

These real world examples show it's one tier away.

Re: Kimi K3: Open Frontier Intelligence

#362
post #111

Pelican: https://tools.simonwillison.net/markdown-svg-renderer#url=ht... - rendered via the OpenRouter API: https://openrouter.ai/moonshotai/kimi-k3 95 input, 16,658 output = 25 cents! https://www.llm-prices.com/#it=95&ot=16658&ic=3&oc=15 (13,241 of those were reasoning tokens.) I think that's the most expensive pelican I've rendered through a Chinese model so far.

Oof, front fork is wrecked. Pelican should be wearing a helmet on that death trap.

[deleted]

Re: Kimi K3: Open Frontier Intelligence

#363

Earlier quoted context omitted.

Right at this moment, there are more people in the world on the side of China than on the side of the USA. Which can translate into raw market numbers at some point. So these comments are kinda moot.

That's not what the actual data shows. The American frontier providers captured the entire market. China is getting the scraps. https://gs.statcounter.com/ai-chatbot-market-share

That is correct, but that’s not what I’m talking about. A lot of people complain about handing their data to Chinese government. My argument is, as of today, people like China more than the US. And the American government has publicly said that they’re basically controlling all AI labs if needed. So yeah.

Re: Kimi K3: Open Frontier Intelligence

#364
post #330

Just in case you were thinking of signing up directly with Moonshot to use the service, they appear to train even on API use: > We may use Content to provide, maintain, develop, support, and improve the Services, comply with applicable law, enforce our terms and policies, and keep the Services safe and secure. Customer who requires restrictions on the use of Customer Content for training or improving Moonshot AI mode…

Interesting. OpenRouter classifies the Moonshot provider as ZDR. I wonder whether they have a ZDR agreement or it's a misclassification on their part.

OpenRouter's ToS also seems to allow them to store your submitted prompts anyway, so privacy advocates would have to look elsewhere anyway, that's at least how I understand it (and it surprised me).

Re: Kimi K3: Open Frontier Intelligence

#365
post #321

Kimi K3 blog is up: https://www.kimi.com/blog/kimi-k3 2.8T param open model, 1M context, native vision. Weights releasing by July 27 with technical report. Launching with max thinking effort by default; low/high effort modes coming in future updates.

These benchmark numbers are insane. The days when China was 6 months behind are over? How are they doing this with so much less resources than the US??? I have so much respect for the researchers there

What makes you think they have less resources?

Re: Kimi K3: Open Frontier Intelligence

#366
post #321

Kimi K3 blog is up: https://www.kimi.com/blog/kimi-k3 2.8T param open model, 1M context, native vision. Weights releasing by July 27 with technical report. Launching with max thinking effort by default; low/high effort modes coming in future updates.

These benchmark numbers are insane. The days when China was 6 months behind are over? How are they doing this with so much less resources than the US??? I have so much respect for the researchers there

I'm not sure where "so much less resources" comes from. Training the best model has nothing to do with having the most NVIDIA GPUs around. If that were true then xAI would have the best model. It comes down to the quality of data, research, and financial backing.

Re: Kimi K3: Open Frontier Intelligence

#367
post #321

Kimi K3 blog is up: https://www.kimi.com/blog/kimi-k3 2.8T param open model, 1M context, native vision. Weights releasing by July 27 with technical report. Launching with max thinking effort by default; low/high effort modes coming in future updates.

These benchmark numbers are insane. The days when China was 6 months behind are over? How are they doing this with so much less resources than the US??? I have so much respect for the researchers there

Mythos/Fable-class models have been around for at least 4 months internally in the US, and Kimi still isn't quite there, so I'd say the 6-months is still about right.

Re: Kimi K3: Open Frontier Intelligence

#368
post #330

Just in case you were thinking of signing up directly with Moonshot to use the service, they appear to train even on API use: > We may use Content to provide, maintain, develop, support, and improve the Services, comply with applicable law, enforce our terms and policies, and keep the Services safe and secure. Customer who requires restrictions on the use of Customer Content for training or improving Moonshot AI mode…

You think openai, anthropic, google, z and any of the others dont? They do, if they say they dont, they do. Who wouldn't in this earth-shattering race. So Naive

Re: Kimi K3: Open Frontier Intelligence

#369
post #177

Earlier quoted context omitted.

I wouldn't be surprised if models were optimizing for rendering SVG pelicans at this point

every ai release thread seems to have this same sequence of comments

My comment on GLM-5 five months ago:

"How many pelican riding bicycle SVGs were there before this test existed? What if the training data is being polluted with all these wonky results..."

https://news.ycombinator.com/item?id=46974853

Re: Kimi K3: Open Frontier Intelligence

#370
post #330

Just in case you were thinking of signing up directly with Moonshot to use the service, they appear to train even on API use: > We may use Content to provide, maintain, develop, support, and improve the Services, comply with applicable law, enforce our terms and policies, and keep the Services safe and secure. Customer who requires restrictions on the use of Customer Content for training or improving Moonshot AI mode…

Interesting. OpenRouter classifies the Moonshot provider as ZDR. I wonder whether they have a ZDR agreement or it's a misclassification on their part.

Why risk it either way if they provide weights for others to run this?

Am I being overly cautious not wanting to send my data to Chinese companies?

Post reply on HN