Live data from Hacker News

The Kimi K3 Moment

stephen.bochinski.dev

1–10 of 644 posts

Re: The Kimi K3 Moment

#6
I think the biggest problem with Chinese models is that they seems to overthink for most of the tasks, especially for smaller ones. The OpenAI models have in my experience only gotten better in terms of efficiency.

Re: The Kimi K3 Moment

#7
> I’ve been running Kimi K3 alongside Claude on my normal coding work, and for all practical purposes I can’t tell them apart

When you say "Claude", do you mean Opus? Fable? What effort level?

Re: The Kimi K3 Moment

#8
I tried Kimi K3 on a task I've done with every other model I use regularly (https://swelljoe.com/post/i-let-every-agent-implement-its-ow...) and found it chewed a lot longer on the problem and ate up almost the entirety of a 5 hour usage limit on their $19 plan.

I only have the $20 plan from OpenAI and the same task, with a lot of the same implementation details as Kimi Code, only took a few minutes and consumed almost none of the 5 hour limit.

Subscription usage limits are hard to measure as none of the providers tell you directly what it means in terms of tokens or anything else you can easily compare, but when I sat down to add Kimi Code to flar, it was because I wanted to try it on some real work and then couldn't do any, because usage was nearly gone after the trivial task...no other ~$20 subscription I have has felt that tight before.

So, it was really slow to complete the task and seemingly much more expensive than every other model I'd tried. Maybe bad luck. Maybe it'll do better on other tasks. I wouldn't know as I was out of usage when I had time to try.

It did find a bug that Gemini 3.5 Flash introduced unprompted, though, so it has that going for it.

Re: The Kimi K3 Moment

#10
post #5
post #3

Earlier quoted context omitted.

considering token efficiency as well I presume?

I'm struggling to decide whether I feel comfortable sending my data to these Chinese models

Are you comfortable sending it to US ones? Especially if installing Claude Code or another tool on your PC and it can collect all the data it wants..

On Openrouter Kimi K3 says it does not retain data or train on it, which is better than what US hosts claim for Claude, ChatGPT, etc.. as they collect and retain data even if you disable training on it.

Opencode or similar open source tool + a zero data retention provider is about the best option aside from running a smaller fully local model on your own PC.

Post reply on HN