Live data from Hacker News

Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

fireworks.ai

41–50 of 491 posts

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#42

I love the Chinese models. I use DeepSeek exclusively and now Kimi K3 offers a great planning assistant for more advanced coding tasks. DeepSeek v4 Flash is extremely fast and is able to handle pretty much anything I've thrown at it (I use mostly Rust, PSQL, Angular and Terraform). I self host Bifrost as my LLM gateway, though I wish LLM vendors would do monthly/daily automatic billing (like VPS providers do) rather…

What don't you like about the service, besides the mark up?

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#43

Why SoTA (uppercase “T”) instead of SotA (lowercase “T”) ? “State of [T]he Art” versus “State of [t]he Art”. If not SotA then at least SOTA, which is more accurate.

[flagged]

Not wrong though. Technically correct and when it comes to communication, also seems like a reasonable query.

In a world of so many acronyms, details matter.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#44

I love the Chinese models. I use DeepSeek exclusively and now Kimi K3 offers a great planning assistant for more advanced coding tasks. DeepSeek v4 Flash is extremely fast and is able to handle pretty much anything I've thrown at it (I use mostly Rust, PSQL, Angular and Terraform). I self host Bifrost as my LLM gateway, though I wish LLM vendors would do monthly/daily automatic billing (like VPS providers do) rather…

Have you looked into OpenCode Zen or OpenCode Go?

https://opencode.ai/zen

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#45
Very interesting. They test Kimi K3 and Fable on a set of approx 1000 tasks grouped into 5 areas (SWE, Legal, etc).

They put a router model in front that predicts whether Kimi or Fable is going to give a better cost for a correct result. (They believe that ultimately such a router model should be continuously trained on your own workloads so it makes the best decisions for you).

Their router chose Kimi the majority of the time (72% in one category, all the way to 96% in another category), leading to cost savings in every category (from 1.5x to 50x depending).

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#46

Earlier quoted context omitted.

You're presenting further evidence of lack of effective control over these systems as... a mitigating factor...?

Why do you assume they're disagreeing with you?

> One man's offensive penetration tool is another mans defensive tool.

seems to suggest the author believes there's some intrinsic equilibrium

Which,

1) is definitely not proven and not guaranteed (open to proofs otherwise, not pithy sayings that have zero normative effect on reality)

2) is apparently "supported by" further evidence of lack of effective control, which does not feel like equilibrium whatsoever

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#47
post #33

Earlier quoted context omitted.

True. However, I believe most non-programmers don't need access to the fanciest model, but just want to use a good LLM without constant nagging about usage limits. Then 19 vs. 20 is true?

Still no. If you're only getting Opus-class you can still end up paying less by just using an equivalent Chinese model on OpenRouter.

Not talking about me or HN users in general. These companies likely need regular peeps to begin using their services in order to become profitable. Not sure these are the ones who will buy from OpenRouter.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#48

Why SoTA (uppercase “T”) instead of SotA (lowercase “T”) ? “State of [T]he Art” versus “State of [t]he Art”. If not SotA then at least SOTA, which is more accurate.

It should be SotA.

Does that make DeepSeek V4 Flash MiniSotA? This dev in the Twin Cities would like to know.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#50
post #14
post #3

A third the cost, open source, and won't refuse every other request because of some vague possible connection to cybersecurity concerns.

My questions about strawberries got blocked as too dangerous! I’m not joking

Strawberries are a well known weakness of LLMs, as they have a hard time to count the numbers of "r"s in them. Maybe that's why, because they fear that weakness could be exploited somehow.
Post reply on HN