Live data from Hacker News

Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

fireworks.ai

131–140 of 491 posts

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#131

we put a router model in front of two other models so the router can decide which model is better at deciding things. next we'll need a router for the router and eventually the entire internet is just routers routing routers to other routers

And the weights are just in the routers now? Love that idea.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#132

Earlier quoted context omitted.

I'm happy that Kimi K3 is indeed SotA and its open weights are due to be released soon. It's also true that Moonshot and other labs distill from Claude. This has been reported on extensively. I don't think there's any alpha for Anthropic distilling from this model. I do not mean to discount the tremendous amount of innovation regarding MoE and quantization that Moonshot has accomplished. But its training with synthet…

When I say distill I also mean mine it architecturally for insights but I doubt seriously the model training is entirely distillation of Claude, it’s almost certainly a mixture of both original corpus and reinforcement as well as distillation. I think it’s a little condescending to imply that these new open models are cheap ripoffs with nothing original to them. These teams and labs are top tier as well, working unde…

I think your characterization of my post, which credits Moonshot's innovation, goes a bit too far.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#133

I have several thoughts about this, which I'll just iterate: 1) US export bans have made it so that Chinese companies have to compete using less-than-state-of-the-art hardware. This has forced Chinese companies to build more cost efficient models. Whereas, US companies have moreso tried to be state of the art by spending more money than anyone else on state-of-the-art hardware. 2) Xi Jinping has called for more open…

[flagged]

Finance mostly true, space launch also true, cancer survival broadly true (with qualifications), oil and natural gas also true.

> (exports nearly 2x the second largest exporter)

Can you provide the source? In the WTO's broader medical goods category, Germany actually exported slightly more than the US in 2022 ($202.6 billion versus $189.6 billion), so such a huge difference in a few years?

> and pharmaceutical research

China accounted for approximately 44.2% of the 104 new molecules in 2025, compared with 26.9% for the U.S. and 15.4% for Europe.

The 2024 shares were approximately 34.6% for China, 30.9% for the U.S., and 22.2% for Europe.

> there is literally not a single major industry where the US is not a leading player.

I will just give 3 examples: China completed roughly 91% of the world’s shipbuilding tonnage in 2025, holds more than 80% of solar module manufacturing capacity and produces more than three quarters of global batteries.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#137

we put a router model in front of two other models so the router can decide which model is better at deciding things. next we'll need a router for the router and eventually the entire internet is just routers routing routers to other routers

And the weights are just in the routers now? Love that idea.

The weights are in the meat

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#138

Why SoTA (uppercase “T”) instead of SotA (lowercase “T”) ? “State of [T]he Art” versus “State of [t]he Art”. If not SotA then at least SOTA, which is more accurate.

[flagged]

I’m overwhelmed by LLM/Agent signal on HN lately. I actually liked this question. It’s trite and whimsical, but frankly nice break.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#140
post #50
post #14

Earlier quoted context omitted.

My questions about strawberries got blocked as too dangerous! I’m not joking

Strawberries are a well known weakness of LLMs, as they have a hard time to count the numbers of "r"s in them. Maybe that's why, because they fear that weakness could be exploited somehow.

Strawberries aren't the weakness, the weakness is the tokenization of a prompt. Any word with multiple duplicate characters is going to be troublesome for LLMs.
Post reply on HN