Anthropic's "durable advantage" theory of US AI dominance is looking pretty silly. There's zero indication that it will be hard for China to keep pace as models improve and start contributing to their own training. Which pretty much invalidates their policy recommendations. They can't even blame it on distillation this time, unless they want to claim that their own preferred security measures were ineffective in prev…
Kimi K3: Open Frontier Intelligence
271–280 of 1001 posts
Re: Kimi K3: Open Frontier Intelligence
#272> Kimi K3 is Kimi’s most capable model to date, with 2.8 trillion parameters. This puts them on the top of the largest open models list: Kimi K3 2.8T DeepSeek-V4-Pro 1.6T (49B active) Kimi K2.6 ~1T (32B active) GLM-5.2 754B (40B active) DeepSeek-V3.2 685B Mistral Large 3 675B That's one mighty large model! Moonshot is going to need the USD 500 million reportedly raised earlier this year to run this model.
I guess it remains to be seen whether this will be open-weights. We don't even know how many active params at this point.
Re: Kimi K3: Open Frontier Intelligence
#273I'm a bit nervous this one isn't going to be open-weights. Any mention of "open" has been struck from the literature for this model (it was present an hour ago). We don't even know active params? At this pricing, I'll be surprised if it's open.
They will release the full weights by 7/27 along with support in vLLM. Source: their release blog on WeChat. https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ
(translated by chrome)
11 days is a long time. It does not take that long to implement inference at providers. In my opinion, seems like they're being pre-emptively cautious about government intervention/review
Re: Kimi K3: Open Frontier Intelligence
#274Earlier quoted context omitted.
Tokenizers also matter. Anthropics tokenizers will encode the same piece of text at a way higher token count than OpenAi, for example. That said, Kimi is competing against GLM in my mind, and GLM 5.2 is less than 1/3 the price.
GLM is actually quite expensive in actual practice because it's not very token efficient. I've yet to find a way to run it on a monthly sub reliably for cheaper than Codex. Neuralwatt was cheap (but slow) but they cranked their price. Ollama monthly sub is speedy but doesn't offer a lot of quota. Right now unless you're paying by the token, there's no cost based reason to use the open weight models for daily coding w…
Re: Kimi K3: Open Frontier Intelligence
#275Why do most LLMs insist on a login, even for a free trial? I entered a question to try it, but as soon as I hit enter it wants my phone number for a login. No thanks.
Re: Kimi K3: Open Frontier Intelligence
#276> Impress me with a 1 page html file
Result: https://ydaurtg3fdwhq.kimi.page/
Came out looking pretty cool! By contrast, Fable produced a moderately more interesting "live observatory" of the solar system.
Re: Kimi K3: Open Frontier Intelligence
#277Earlier quoted context omitted.
It's like reading Anthropic's obituary.
This is weird and reactionary. Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns. Anthropic/american models aren't going anywhere anytime soon.
With Oracle being junk before this, more will follow.
Re: Kimi K3: Open Frontier Intelligence
#278Earlier quoted context omitted.
It's like reading Anthropic's obituary.
This is weird and reactionary. Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns. Anthropic/american models aren't going anywhere anytime soon.
Re: Kimi K3: Open Frontier Intelligence
#279Re: Kimi K3: Open Frontier Intelligence
#280Earlier quoted context omitted.
It's like reading Anthropic's obituary.
This is weird and reactionary. Lots of organizations are continuing to refuse to use chinese models due to security and IP concerns. Anthropic/american models aren't going anywhere anytime soon.