Live data from Hacker News

Kimi K3: Open Frontier Intelligence

kimi.com

241–250 of 1001 posts

Re: Kimi K3: Open Frontier Intelligence

#241

I'm a bit nervous this one isn't going to be open-weights. Any mention of "open" has been struck from the literature for this model (it was present an hour ago). We don't even know active params? At this pricing, I'll be surprised if it's open.

[deleted]

Re: Kimi K3: Open Frontier Intelligence

#242
post #228

Imagine you're a mid sized company and you can host this model locally. Suddenly there are zero reasons to pay a single red cent to the bloodsucking American AI cartel.

Whether it is "open" or not seems to be in question. While it was initially called an "open" model, it seems that "open" mentions have been scrubbed from website.

Re: Kimi K3: Open Frontier Intelligence

#243
post #227

Some official benchmark numbers posted in Chinese social media (I am sure they will publish an English blogpost later too): https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ Generally looks like a Sol/Fable tier model, better across the board than Opus 4.8. (Edit) English blogpost is up now: https://www.kimi.com/blog/kimi-k3

It's like reading Anthropic's obituary.

Re: Kimi K3: Open Frontier Intelligence

#244
post #50
post #37

> In our evaluations, Kimi K3 delivers frontier-level performance. Among the models tested, its overall intelligence ranks second only to Claude Fable 5 and GPT-5.6 Sol. For the complete benchmark results, see our tech blog. The full model weights of Kimi K3 will be released in the coming days. More details on the architecture, training, and evaluation will be published together with the Kimi K3 technical report. > K…

> its overall intelligence ranks second only to Claude Fable 5 and GPT-5.6 Sol Pretty sure ranking “second” to two others means ranking third.

Not if the others tie for first place.

Re: Kimi K3: Open Frontier Intelligence

#245
post #228

Imagine you're a mid sized company and you can host this model locally. Suddenly there are zero reasons to pay a single red cent to the bloodsucking American AI cartel.

Can you host the model for a lower cost per token than you'd pay Anthropic or OpenAI for a similar level of intelligence? I doubt you're beating their efficiencies of scale.

Re: Kimi K3: Open Frontier Intelligence

#246

I'm a bit nervous this one isn't going to be open-weights. Any mention of "open" has been struck from the literature for this model (it was present an hour ago). We don't even know active params? At this pricing, I'll be surprised if it's open.

They will release the full weights by 7/27 along with support in vLLM.

Source: their release blog on WeChat. https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ

Re: Kimi K3: Open Frontier Intelligence

#247

The big danger here is the gradual increase in open-weight subscription costs. I use open weight subscriptions, with lower-cost models for 80% of my tasks and GLM-5.2, Qwen 3.7-Max, Kimi-K2.6/2.7-Code for the 20% that need the most intelligence. That lets me maximize the rate-limit the subscription gives (rate limits per model are literally a price-limit-per-token/model). When new/more expensive open weights come in,…

It’s open weight, so the price will end up being the marginal cost of hosting it.

Personally, I like that there is an option to not send data to companies that have strong financial incentives to steal it.

Also, open weight foundation models can be distilled, so they’re providing a service that the US duopoly is actively blocking. Given that app specific distillation can get > 10x improvements on inference cost (with slight improvement of quality), it’s clear that it’ll win out over time.

Re: Kimi K3: Open Frontier Intelligence

#248

Curious why the thinking mention chatgpt for a moment https://ibb.co/JFdhMN95

LLMs are hopelessly confused about which model they are. Ask DeepSeek V4 Flash which model it is, and it's 50/50 between "I am DeepSeek (深度求索)" and "I am part of the GPT-4 series developed by OpenAI." Ask Claude, it'll say Claude. Ask Claude in Chinese, it'll sometimes say DeepSeek.

It's incredibly funny, but I don't know whether it's related to distillation; it's probably quite rare for a distilled trace to mention which model it came from. (I'm not saying distillation doesn't happen, just that it's possibly unrelated.)

For your specific example, the internet is full of "As a large language model developed by OpenAI, I can't..." due to people pasting chatbot output without reading it. Seems reasonable for that to surface as part of the CoT for your question about model capabilities.

Re: Kimi K3: Open Frontier Intelligence

#249
post #227

Some official benchmark numbers posted in Chinese social media (I am sure they will publish an English blogpost later too): https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ Generally looks like a Sol/Fable tier model, better across the board than Opus 4.8. (Edit) English blogpost is up now: https://www.kimi.com/blog/kimi-k3

It's like reading Anthropic's obituary.

Fable is by Anthropic, and this is too expensive, GLM 5.2 is roughly the same quality at a much cheaper price.

(I mantain a client with llama.cpp and 101 models across 14 companies by http)

Re: Kimi K3: Open Frontier Intelligence

#250
post #177

Earlier quoted context omitted.

I wouldn't be surprised if models were optimizing for rendering SVG pelicans at this point

every ai release thread seems to have this same sequence of comments

I wouldn't be surprised if models were optimizing for pelican-related comment chains at this point
Post reply on HN