Live data from Hacker News

Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

fireworks.ai

91–100 of 491 posts

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#92

Earlier quoted context omitted.

If weights are the source for models then ELF binaries are the source for software.

clearly not true. the weights are the preferred form for making modifications. Do you really think people should be downloading hundreds of TB of training data and running make to build the model on their own cluster of GPUs?

[deleted]

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#93
post #75

Earlier quoted context omitted.

The best part is using harnesses like reasonix or whale make cache hit at a rate close to 98%, making requests converge to practically free. And that's with unsubsidized American providers like cloudflare or Digital Ocean.

How can you be hitting cache on what I think are novel LLM prompts …

Not them but my understanding is that the harness will send a simple 'heartbeat' message to keep the cache 'warm', (see prefix caching: https://handbook.modular.com/inference-optimization/prefix-c... ) which can then be edited/changed, which does cause the user to incur a fee, but its much less than the amount they'd pay on a no-cache hit request.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#94

[flagged]

I'm not sure why you're so dismissive about this model. It isn't currently open, but it will be. It was slow when I used it so i'll give you that. There will be other providers with better performance.

My only gripe was that it takes a very long time to think. It's a good model otherwise.

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#95
post #75

Earlier quoted context omitted.

The best part is using harnesses like reasonix or whale make cache hit at a rate close to 98%, making requests converge to practically free. And that's with unsubsidized American providers like cloudflare or Digital Ocean.

How can you be hitting cache on what I think are novel LLM prompts …

multi turn sessions, they are typically in the high 90% hit rate across all providers without doing much of anything

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#97

we put a router model in front of two other models so the router can decide which model is better at deciding things. next we'll need a router for the router and eventually the entire internet is just routers routing routers to other routers

[dead]

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#98
post #88

Earlier quoted context omitted.

They share their methodology and results. I learned things about the relative strengths and weaknesses of Kimi and Fable I hadn’t seen anywhere else. Should being in the model hosting business disqualify them from sharing?

Doesn't disqualify them, but it may call into question their results seeing as they have a potential conflict of interest.

[deleted]

Re: Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA

#99

Earlier quoted context omitted.

If weights are the source for models then ELF binaries are the source for software.

clearly not true. the weights are the preferred form for making modifications. Do you really think people should be downloading hundreds of TB of training data and running make to build the model on their own cluster of GPUs?

Random people? No. Governments and big corporations? Yes. It removes any concern of "backdoors", and is currently the best starting point for your own model which will be as capable as k3.
Post reply on HN