Live data from Hacker News

Kimi K3: Open Frontier Intelligence

kimi.com

951–960 of 1001 posts

Re: Kimi K3: Open Frontier Intelligence

#951

This might be the most impressive website generator demo I've seen: https://macos27.kimi.page Context from the person who prompted it: https://x.com/mweinbach/status/2077827886149439547

I'm pretty annoyed with how fast this feels. Wish MacOS was this fast launching things.

[deleted]

Re: Kimi K3: Open Frontier Intelligence

#952
post #629

Earlier quoted context omitted.

> Companies can still make money from commodities Especially Chinese companies. Just think about all the other industries where Chinese companies dominate by extremely low cost.

Isn't that typically because they have lower labor costs? Not sure that applies as much to AI.

The part that applies to AI is probably that they're (accused of) distilling SOTA models at a massive scale.

Re: Kimi K3: Open Frontier Intelligence

#953

This might be the most impressive website generator demo I've seen: https://macos27.kimi.page Context from the person who prompted it: https://x.com/mweinbach/status/2077827886149439547

It looks better and the fonts are more readable than the native wayland/gnome display I'm viewing it on.

Re: Kimi K3: Open Frontier Intelligence

#954
post #518

Earlier quoted context omitted.

If there was some grand strategy for all Chinese labs, surely it'd have leaked by now. I think its more likely that: - Companies can still make money from commodities - Chinese labs only have 5-10% the valuation of OpenAI/Anthropic, so massive monopoly profits aren't necessary. Profit expectations for tech companies in China are really low in general, complete opposite of the US. - Open weighting is a great way to ge…

> Chinese labs only have 5-10% the valuation of OpenAI/Anthropic, so massive monopoly profits aren't necessary. Profit expectations for tech companies in China are really low in general, complete opposite of the US. It’s really amazing to see that the competition is creating better quality models for everyone - and am really happy that some of these are open source (or partially os). Regarding the valuation, that may…

[deleted]

Re: Kimi K3: Open Frontier Intelligence

#955
I just tried this on the monthly $18 plan, having it do a basic task with its 2.7 model and then audit it using k3.

K3 got into some loop trying to run docker and after maybe the 6th attempt ran out of quota for the 5 hour window which represents 20% of the weekly.

I run 200 max and chatgpt pro, but I had to blink at that.

K3 didn't even write out what it was doing or provide any sense for why it was pursuing the execution path it was.

I'm in disbelief that this is a groundbreaking model, and do not think it represents a threat to Claude Code or Codex at this time.

Re: Kimi K3: Open Frontier Intelligence

#956

Earlier quoted context omitted.

Hey Simon, I noticed one thing all LLMs are currently pretty bad at and maybe we could create a benchmark from it. Let an LLM play the role of a dungeon master and tell it to strictly stay in the script/story and only allow realistic player actions. You will notice that they are easily brought off track. E.g. - Tell the LLM that you as a player noticed a strange glow in an NPCs eyes -> the NPC becomes an enemy. - In…

This is a known and solved problem. Such a test is pointless for a general-purpose model, because like most people you're using multiturn chats in a naive way, fighting the default finetuning that is done intentionally. 1. You're sending your in-character inputs to an instruction-tuned model under the user role, in a multiturn chat. It's biased to treat these inputs as instructions and this behavior will show itself…

Are you testing a model or a harness? People conflate the two.

Re: Kimi K3: Open Frontier Intelligence

#957
post #840

This might be the most impressive website generator demo I've seen: https://macos27.kimi.page Context from the person who prompted it: https://x.com/mweinbach/status/2077827886149439547

Even the terminal works, this is madness. I was able to create a temp folder, echo hello > world, and then I could open the folder in finder and double-clicking the file opens it in a GUI text editor.

This is wild. Do you think it's really a one prompter?

K2.5 had a linux frontend one-shot for display that was very good looking and smooth but very little of it had function. Should I just like, idk, stop using subscriptions and API this shiz?

Re: Kimi K3: Open Frontier Intelligence

#958

Earlier quoted context omitted.

Distillation is not an attack. It simply a way to train a model. Not doing it when you are behind is akin to snatching defeat from the jaws of victory.

How Xi coded. If you just cheat off the top student in your class that's simply a way to get a good GPA.

Ah yes, because if a person agrees with anything a chinese company does it must be because they love Xi. Get real.

Munching of the top student in a class is clearly prohibited. Distillation is not cheating, it's learning from your competitor. Akin to a company purchasing their competitor thingamajig to see if you can improve their own product.

Re: Kimi K3: Open Frontier Intelligence

#959

This might be the most impressive website generator demo I've seen: https://macos27.kimi.page Context from the person who prompted it: https://x.com/mweinbach/status/2077827886149439547

I just favorited your comment, which joins a vanishingly small list of revelations which completely blew my mind

It would be so cool as a feature to allow people on HN to make some of their favorited posts/comments public.

Re: Kimi K3: Open Frontier Intelligence

#960

I just tried this on the monthly $18 plan, having it do a basic task with its 2.7 model and then audit it using k3. K3 got into some loop trying to run docker and after maybe the 6th attempt ran out of quota for the 5 hour window which represents 20% of the weekly. I run 200 max and chatgpt pro, but I had to blink at that. K3 didn't even write out what it was doing or provide any sense for why it was pursuing the exe…

Have you used the same session for audit? so switched to K3? or used a new session for K3? K3 is sensitive to this, they wrote about it on their blog.
Post reply on HN