Live data from Hacker News

Kimi K3: Open Frontier Intelligence

kimi.com

271–280 of 1001 posts

Re: Kimi K3: Open Frontier Intelligence

#271

Anthropic's "durable advantage" theory of US AI dominance is looking pretty silly. There's zero indication that it will be hard for China to keep pace as models improve and start contributing to their own training. Which pretty much invalidates their policy recommendations. They can't even blame it on distillation this time, unless they want to claim that their own preferred security measures were ineffective in prev…

Likely won't improve much. They trained on every text already.

Re: Kimi K3: Open Frontier Intelligence

#272
post #94
post #86

> Kimi K3 is Kimi’s most capable model to date, with 2.8 trillion parameters. This puts them on the top of the largest open models list: Kimi K3 2.8T DeepSeek-V4-Pro 1.6T (49B active) Kimi K2.6 ~1T (32B active) GLM-5.2 754B (40B active) DeepSeek-V3.2 685B Mistral Large 3 675B That's one mighty large model! Moonshot is going to need the USD 500 million reportedly raised earlier this year to run this model.

I guess it remains to be seen whether this will be open-weights. We don't even know how many active params at this point.

[deleted]

Re: Kimi K3: Open Frontier Intelligence

#273
post #246

I'm a bit nervous this one isn't going to be open-weights. Any mention of "open" has been struck from the literature for this model (it was present an hour ago). We don't even know active params? At this pricing, I'll be surprised if it's open.

They will release the full weights by 7/27 along with support in vLLM. Source: their release blog on WeChat. https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ

>We are currently working closely with our inference partners and open-source maintainers to align the technical details and ensure the model can be reliably deployed across the ecosystem. The full model weights will be released by July 27, 2026. Further details regarding the architecture, training, and evaluation will be released with the Kimi K3 technical report.

(translated by chrome)

11 days is a long time. It does not take that long to implement inference at providers. In my opinion, seems like they're being pre-emptively cautious about government intervention/review

Re: Kimi K3: Open Frontier Intelligence

#274

Earlier quoted context omitted.

Tokenizers also matter. Anthropics tokenizers will encode the same piece of text at a way higher token count than OpenAi, for example. That said, Kimi is competing against GLM in my mind, and GLM 5.2 is less than 1/3 the price.

GLM is actually quite expensive in actual practice because it's not very token efficient. I've yet to find a way to run it on a monthly sub reliably for cheaper than Codex. Neuralwatt was cheap (but slow) but they cranked their price. Ollama monthly sub is speedy but doesn't offer a lot of quota. Right now unless you're paying by the token, there's no cost based reason to use the open weight models for daily coding w…

I'm on the Z.ai quarterly subscription plan (got in when the price was lower) and I was using it through opencode and it was like I'd only get maybe an hour of usage (if that, sometimes) before it would time out and say come back in 5 hours. Now I'm using it through their Zcode harness and I rarely hit that - they say they're giving 1.5x usage if you use it through Zcode, sometimes seems like even more than that.

Re: Kimi K3: Open Frontier Intelligence

#275

Why do most LLMs insist on a login, even for a free trial? I entered a question to try it, but as soon as I hit enter it wants my phone number for a login. No thanks.

Free use without registration -> free to anyone and anything -> easy to abuse at scale, with no way to restrict use.

Re: Kimi K3: Open Frontier Intelligence

#276
My testing prompt for these models is by no means objective or repeatable (like the pelican) but it's a nice test of curiosity:

> Impress me with a 1 page html file

Result: https://ydaurtg3fdwhq.kimi.page/

Came out looking pretty cool! By contrast, Fable produced a moderately more interesting "live observatory" of the solar system.

Re: Kimi K3: Open Frontier Intelligence

#277

Earlier quoted context omitted.

It's like reading Anthropic's obituary.

This is weird and reactionary. Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns. Anthropic/american models aren't going anywhere anytime soon.

Nope, but I think this is maybe the critical mass needed to finally crash the AI hype/datacenter cost problem everyones is talking about.

With Oracle being junk before this, more will follow.

Re: Kimi K3: Open Frontier Intelligence

#278

Earlier quoted context omitted.

It's like reading Anthropic's obituary.

This is weird and reactionary. Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns. Anthropic/american models aren't going anywhere anytime soon.

You can run open weight models anywhere.

Re: Kimi K3: Open Frontier Intelligence

#280

Earlier quoted context omitted.

It's like reading Anthropic's obituary.

This is weird and reactionary. Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns. Anthropic/american models aren't going anywhere anytime soon.

If it ends up being open weights, companies will use it running in US data centers.
Post reply on HN