The Kimi K3 Moment
stephen.bochinski.dev
The Kimi K3 Moment
1–10 of 644 posts
Re: The Kimi K3 Moment
#2Re: The Kimi K3 Moment
#3Half-OT: can anyone recommend a LLM cost calculator that's up to date?
Re: The Kimi K3 Moment
#4Re: The Kimi K3 Moment
#5Re: The Kimi K3 Moment
#6Re: The Kimi K3 Moment
#7When you say "Claude", do you mean Opus? Fable? What effort level?
Re: The Kimi K3 Moment
#8I only have the $20 plan from OpenAI and the same task, with a lot of the same implementation details as Kimi Code, only took a few minutes and consumed almost none of the 5 hour limit.
Subscription usage limits are hard to measure as none of the providers tell you directly what it means in terms of tokens or anything else you can easily compare, but when I sat down to add Kimi Code to flar, it was because I wanted to try it on some real work and then couldn't do any, because usage was nearly gone after the trivial task...no other ~$20 subscription I have has felt that tight before.
So, it was really slow to complete the task and seemingly much more expensive than every other model I'd tried. Maybe bad luck. Maybe it'll do better on other tasks. I wouldn't know as I was out of usage when I had time to try.
It did find a bug that Gemini 3.5 Flash introduced unprompted, though, so it has that going for it.
Re: The Kimi K3 Moment
#9Half-OT: can anyone recommend a LLM cost calculator that's up to date?
Re: The Kimi K3 Moment
#10Earlier quoted context omitted.
considering token efficiency as well I presume?
I'm struggling to decide whether I feel comfortable sending my data to these Chinese models
On Openrouter Kimi K3 says it does not retain data or train on it, which is better than what US hosts claim for Claude, ChatGPT, etc.. as they collect and retain data even if you disable training on it.
Opencode or similar open source tool + a zero data retention provider is about the best option aside from running a smaller fully local model on your own PC.