https://www.kimi.com/blog/kimi-k3
Maybe we should update the link to it instead?
351–360 of 1001 posts
https://www.kimi.com/blog/kimi-k3
Maybe we should update the link to it instead?
Earlier quoted context omitted.
GLM is actually quite expensive in actual practice because it's not very token efficient. I've yet to find a way to run it on a monthly sub reliably for cheaper than Codex. Neuralwatt was cheap (but slow) but they cranked their price. Ollama monthly sub is speedy but doesn't offer a lot of quota. Right now unless you're paying by the token, there's no cost based reason to use the open weight models for daily coding w…
I know GLM is relatively expensive and so is Kimi, in comparison to those DeepSeek V4 pro and flash are a godsend and are absolutely good value.
The big danger here is the gradual increase in open-weight subscription costs. I use open weight subscriptions, with lower-cost models for 80% of my tasks and GLM-5.2, Qwen 3.7-Max, Kimi-K2.6/2.7-Code for the 20% that need the most intelligence. That lets me maximize the rate-limit the subscription gives (rate limits per model are literally a price-limit-per-token/model). When new/more expensive open weights come in,…
It’s open weight, so the price will end up being the marginal cost of hosting it. Personally, I like that there is an option to not send data to companies that have strong financial incentives to steal it. Also, open weight foundation models can be distilled, so they’re providing a service that the US duopoly is actively blocking. Given that app specific distillation can get > 10x improvements on inference cost (with…
Wonder if they’ll open-source this and show how many tokens it cost.
My testing prompt for these models is by no means objective or repeatable (like the pelican) but it's a nice test of curiosity: > Impress me with a 1 page html file Result: https://ydaurtg3fdwhq.kimi.page/ Came out looking pretty cool! By contrast, Fable produced a moderately more interesting "live observatory" of the solar system.
My testing prompt for these models is by no means objective or repeatable (like the pelican) but it's a nice test of curiosity: > Impress me with a 1 page html file Result: https://ydaurtg3fdwhq.kimi.page/ Came out looking pretty cool! By contrast, Fable produced a moderately more interesting "live observatory" of the solar system.
Earlier quoted context omitted.
I've been avidly using Fable since it was re-released and while it has been excellent at building the apps I want, the reasoning has been completely opaque. Kim, however, has exposed the whole reasoning trace, or enough of it to matter. I'd almost forgotten how nice it is to see this. I've been able to see all of the weird twist and turns it takes and it is joyful. But also, far, far more informative and means I can…
I have severe complaints about Anthropic's product managers on this front. Their preference for hiding, obscuring, and trying to wrest control from the user are a bit harrowing. It would be wonderful to go back to Claude Code from before March. It seems like every release destroys value for me!
Say of that what you will, but it's not because they want to wrest control from users.
It's because they don't want Chinese companies to do exactly what Moonshot (Kimi creators) and others have done.
Just in case you were thinking of signing up directly with Moonshot to use the service, they appear to train even on API use: > We may use Content to provide, maintain, develop, support, and improve the Services, comply with applicable law, enforce our terms and policies, and keep the Services safe and secure. Customer who requires restrictions on the use of Customer Content for training or improving Moonshot AI mode…
Interesting. OpenRouter classifies the Moonshot provider as ZDR. I wonder whether they have a ZDR agreement or it's a misclassification on their part.
That said, I wouldn't rule out OpenRouter misclassifying - I've seen some providers where I'm fairly sure they have.
Some official benchmark numbers posted in Chinese social media (I am sure they will publish an English blogpost later too): https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ Generally looks like a Sol/Fable tier model, better across the board than Opus 4.8. (Edit) English blogpost is up now: https://www.kimi.com/blog/kimi-k3
Earlier quoted context omitted.
Right at this moment, there are more people in the world on the side of China than on the side of the USA. Which can translate into raw market numbers at some point. So these comments are kinda moot.
Maybe the Democracy Index can make this a little more fact-based: https://en.wikipedia.org/wiki/The_Economist_Democracy_Index USA = Flawed democracy China = Authoritarian I don't really know how well they do this index, but probably better than a random HN comment.