Live data from Hacker News

Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

fireworks.ai

221–230 of 491 posts

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#221
post #115

What's the data governance and privacy controls on using Kimi K3 if I subscribe to their coding plans? I want to migrate away from Anthropic

Need to wait until "western" providers start hosting it.

Fireworks (the author of OP's article) is a western provider based in San Mateo, California

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#225
post #218

I will accept a 5% drop in benchmarks for a model that talks to me like a human.

I strictly prefer when models ignore any human quirks in my responses. Claude trying to be your friend, saying LOL to your jokes is ridiculous and frankly, harmful

Maybe we're prompting it different, but it's not "trying to be my friend" for sure, nor am I trying to be "its" friend either. Or at least I'm sufficiently oblivious to its advances, and find it unthinkable to form such a bond :)

On the flipside, it does spuriously make hilarious remarks like "Good data.", which I find pretty funny specifically because it comes across as just silly. Not sure how it'd be harmful either, a little entertainment I think goes a long way in this type of profession.

I see zero issues with these, and I have a hard time understanding why people have their panties in a twist so hard about them. I sometimes really quite wonder just what kind of correspondence would y'all prefer, and how would that sound like.

Matter of fact, do you have an example at hand? Like an exact before & after?

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#226
post #178

I love the Chinese models. I use DeepSeek exclusively and now Kimi K3 offers a great planning assistant for more advanced coding tasks. DeepSeek v4 Flash is extremely fast and is able to handle pretty much anything I've thrown at it (I use mostly Rust, PSQL, Angular and Terraform). I self host Bifrost as my LLM gateway, though I wish LLM vendors would do monthly/daily automatic billing (like VPS providers do) rather…

Yeah deepseek is my workhorse of choice these days, I mostly use pro because I do a lot of concurrent jobs so speed is less of a concern. When it falls over I change the model mid-chat to GPT5.5 and chuck a couple of tokens over to OpenAI and then once it's correctly found the problem I switch back to deepseek, keeping all the 5.5 analysis in context. I have been a heavy user of k2.5/6 but they seem to have gotten sl…

> GLM 5.2 has been good at select tasks but really shit at others

Yeah it really likes to shit the bed after thinking forever. Even if it can do some things, the quality is not quite up there in my experience.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#228
post #16

Anyone have routing harnesses like this describes with Claude Code? Or other good routing platform recommendations? (yes, I know this article is about an oracle router)

There's this https://github.com/code-yeongyu/oh-my-openagent which implements the OP article's oracle pattern across 11 roles. Each role has a whole ranking of recommended LLMs across many providers. For example "Sisyphus (claude-opus-4-8 / kimi-k3 / glm-5 ) is your main orchestrator."

Also, you can't use your claude subscription with oh-my-openagent. But you can with Kimi. ALso K3 is on OpenCode GO right now (low limits, but it's possible)

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#229
post #115

Earlier quoted context omitted.

Need to wait until "western" providers start hosting it.

Fireworks (the author of OP's article) is a western provider based in San Mateo, California

Fireworks isn't serving Kimi K3 yet. Presumably, they ran this benchmark against the Moonshot API.

All of the Western providers with sufficient capacity will be able to make it available when the weights are released Monday.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#230

I love the Chinese models. I use DeepSeek exclusively and now Kimi K3 offers a great planning assistant for more advanced coding tasks. DeepSeek v4 Flash is extremely fast and is able to handle pretty much anything I've thrown at it (I use mostly Rust, PSQL, Angular and Terraform). I self host Bifrost as my LLM gateway, though I wish LLM vendors would do monthly/daily automatic billing (like VPS providers do) rather…

I didn't realize that OpenRouter had a mark up. Is it a flat mark up across the board or depending on the model?
Post reply on HN