Live data from Hacker News

Kimi K3: Open Frontier Intelligence

kimi.com

321–330 of 1001 posts

Re: Kimi K3: Open Frontier Intelligence

#321

Kimi K3 blog is up: https://www.kimi.com/blog/kimi-k3 2.8T param open model, 1M context, native vision. Weights releasing by July 27 with technical report. Launching with max thinking effort by default; low/high effort modes coming in future updates.

These benchmark numbers are insane. The days when China was 6 months behind are over? How are they doing this with so much less resources than the US??? I have so much respect for the researchers there

Re: Kimi K3: Open Frontier Intelligence

#322

Earlier quoted context omitted.

This is weird and reactionary. Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns. Anthropic/american models aren't going anywhere anytime soon.

> Lots of organizations are continuing to refuse to use chinese models Correction: Lots of organizations are refusing to use Anthropic Fable because they have forced opt-in data collection as part of their privacy policy, even for Enterprise.

Both things, and both reasons, can be true at the same time.

Not everyone's going to care about Anthropic requiring data collection (a similar debate plays out with regards to "pay or consent" on website tracking), just as not everyone cares about China with regards to security/IP issues (if they did, a lot more would be banned besides occasionally-Huawei).

Re: Kimi K3: Open Frontier Intelligence

#323
post #315

Earlier quoted context omitted.

Free use without registration -> free to anyone and anything -> easy to abuse at scale, with no way to restrict use.

You can limit it a lot to minimize the abuse. In free entrypoint, set token and context limits to be very small. Limit to 2 prompts per IP or something every X hour. That is already a substantial limit where bypassing might not provide much benefits.

Residential proxies are too prevalent for IP address limits to work effectively.

Re: Kimi K3: Open Frontier Intelligence

#324

Another deepseek moment? it seems they have fully caught with fable tier of models, and this was a lot sooner than was expected.

Yeah, I would have expected Zhipu to ship a Fable-adjacent model by the end of the year, but the jump from Kimi 2.7 (which I think is just barely at the level where it is genuinely helpful for coding) to this is absolutely bonkers. And this is clearly not just benchmaxing; this thing actually works. If you told me I could only use this and never use Fable or Sol again, I'd shrug and not feel like I'd lost much.

> Yeah, I would have expected Zhipu to ship a Fable-adjacent model by the end of the year

There were talks of a GLM 5.3 in August, so maybe not that far away...

Re: Kimi K3: Open Frontier Intelligence

#325
post #312

Earlier quoted context omitted.

This is weird and reactionary. Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns. Anthropic/american models aren't going anywhere anytime soon.

> Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns This is such a common omission: the Chinese models are open, you can host them yourself on your premises. So privacy and independence.

it's well documented that models can be adversarially trained with essentially backdoors in response to special inputs

while I am skeptical that this is happening atm, there are probably many industries where the risk does not seem worthwhile

Re: Kimi K3: Open Frontier Intelligence

#326

Kimi doesn't do well on my "ask a trivia question that other AIs get wrong" test. The question it came up with, "which U.S. state is closest to Africa?" is a pretty standard trivia question without any reason to believe other AIs would get confused. https://pellmell.ai/s/dccdeca69f929f79bc89317035610049 Even GPT-OSS-120b gets this right: https://pellmell.ai/s/1a43dfc7a3baa214aa0fa1b95d2c536a

Are you giving it your API for these other AIs to evaluate their responses? This 'test' seems perverse.

Re: Kimi K3: Open Frontier Intelligence

#327

Working with chinese models is giving me a fullfilment sensation. I think that I have enough quality for the work that I need to do and lots of extra tokens to work with. With Claude and ChatGPT I reach the limits fairly easy, but not with OpenCode Go. So I will use Claude once in a while for difficult tasks to see how much better it still is (but use Chinese on a daily basis)

I have been using Deepseek V4 Pro for personal projects and it has been great. I think the $20/mo GPT plan is still the strongest value, but only because you don’t have to pay API prices for tokens.

Re: Kimi K3: Open Frontier Intelligence

#328

Anthropic's "durable advantage" theory of US AI dominance is looking pretty silly. There's zero indication that it will be hard for China to keep pace as models improve and start contributing to their own training. Which pretty much invalidates their policy recommendations. They can't even blame it on distillation this time, unless they want to claim that their own preferred security measures were ineffective in prev…

Likely won't improve much. They trained on every text already.

most of the gains from the past year and a half have not been from web data, but from synthetic data and agent rollouts with RL.

Re: Kimi K3: Open Frontier Intelligence

#329
post #255

Earlier quoted context omitted.

> Maybe another DeepSeek moment right here. Surely not... What made DeepSeek disruptive was that the cost was 10X lower. In this case, the cost is about 2X lower the Sol I think? At 2X, you're pretty close to the error margins due to token efficiency etc... I'd say this is "on trend" for open models catching up to frontier labs, but its not a "change in the trend" like DeepSeek was IMO.

DeepSeek didn’t really change any trends though, unless you count the stock market. It was impressive work, but models were commoditizing and inference costs were dropping rapidly already. They were neither the first nor the last 10x optimization, from what I’ve seen.

If you know of any other 10x optimisations currently, please let me know! I'm in the market for a model that's a tenth the price of a frontier model at the same level of quality.

Re: Kimi K3: Open Frontier Intelligence

#330
Just in case you were thinking of signing up directly with Moonshot to use the service, they appear to train even on API use:

> We may use Content to provide, maintain, develop, support, and improve the Services, comply with applicable law, enforce our terms and policies, and keep the Services safe and secure. Customer who requires restrictions on the use of Customer Content for training or improving Moonshot AI models may contact Moonshot AI to discuss available enterprise arrangements or separate written agreements. Unless otherwise expressly agreed in writing, Customer Content may be used for the foregoing purposes.

https://platform.kimi.ai/docs/agreement/modeluse#4-content

Post reply on HN