Live data from Hacker News

The Kimi K3 Moment

stephen.bochinski.dev

151–160 of 644 posts

Re: The Kimi K3 Moment

#151

Earlier quoted context omitted.

> western governments Are you talking about the US, specifically? Why would other countries, that don't share the same anxiety about China as the US, would be troubled with the this?

The US might pressure them?

Not sure if you noticed, European countries are distancing themselves from the US. They couldn't be pressured to offer logistic support to the US shitshow in Iran, why would they be pressured to help the US in its protectionism of its AI bubble?

Re: The Kimi K3 Moment

#152
post #112

Earlier quoted context omitted.

I was more thinking they would be funding US labs.

the question was: what is the endgame for the stated "second class labs" strategy of distilling their frontier competitors then undercutting them on price?

Making lots of money?

Re: The Kimi K3 Moment

#153

This was always where this was heading, but we got here much faster than expected. Once western governments declare it to be a "national security" risk for citizens to have access to open-weight frontier models, and once they classify using these models as acts of terrorism, what will that world be like? Will using Kimi K3 come to be like how napster was in the olden days? Everybody knew it was technically illegal, b…

> western governments Are you talking about the US, specifically? Why would other countries, that don't share the same anxiety about China as the US, would be troubled with the this?

Well, many EU countries like Italy and Germany officially freaked out about DeepSeek, ordering it to be banned from app stores etc.

Re: The Kimi K3 Moment

#154
post #146
post #29

Earlier quoted context omitted.

The visa that would correlate to this is the O-1 visa 20k O-1 visas were issued last FY which was mostly under the Trump admin, up from 19.5k the previous FY under the Biden admin

The O-1 has also been abused for a long time, basically any software engineer kid who gets into Y Combinator has been getting an O-1

[deleted]

Re: The Kimi K3 Moment

#155
post #116

Earlier quoted context omitted.

At this point, the United States will lose that battle most Countries in the world are going end up using electronics from Asia, that ship has sailed Japan, China, Korea, Singapore, Taiwan, Vietnam, dominate that area, China already dominates EVs, Drones and many other electronic devices, and with the way Donald Trump has picked fights, Europe, Canada, Australia, New Zealand, Mexico and many others are looking for ot…

[flagged]

Hilarious response to concerns about industrial production, thanks for getting us here.

Re: The Kimi K3 Moment

#156

I never truly understood what the intended business model around LLMs was. Get them widespread through cheap pricing and then jacking it up? Being the only ones that had a viable product so to get the ability to extract as much value as you want from AI? I don't understand how a product that: - is interfaced with and is deeply linked to natural language, so everything you produce (sessions, history, etc) is in Markdo…

The valuation is based on one lab getting a decisive first advantage, and turning that into a durable self-improving advantage that can never be caught up to. If any can pull it off (a gigantic if), they will effectively own most AI value, and the people who own their shares will live happily ever after. Divide your investment between the labs that could plausibly do this, and your EV may not be dreadful.

And then you wake up from the dream…

Re: The Kimi K3 Moment

#157

Earlier quoted context omitted.

This line of thinking makes no sense because it assumes that labs that distill from frontier models are doing nothing else. It's the classic "the Chinese can only copy" mentality, and it's going to end poorly for American companies. I'm pretty sure that all labs are distilling each others' LLMs, maybe apart from Anthropic and OpenAI. It would be stupid not to do it, because it's cheap and effective. But that's not th…

i never assumed that, and i do keep up with the publications. i'm also not saying it's a dumb thing to do! what i am saying is that empirically, it appears that distillation of a more advanced model is a required first step for them to train a borderline competitive, cheaper model. in effect, their training is subsidized by the frontier labs. if this were not the case, then we would be observing chinese models that f…

> empirically, it appears that distillation of a more advanced model is a required first step

I see no evidence for that.

> if this were not the case, then we would be observing chinese models that far surpass frontier models

It's pretty clear that the primary reason for the difference is budget and compute availability. Chinese labs have at least an order of magnitude less money than Anthropic and OpenAI.

> what happens to these efforts when the subsidy is cut off?

They will continue making progress as they do now, minus the benefits of distillation.

Re: The Kimi K3 Moment

#158
A lot of these open source models do look good on public benchmark but not sure if they are that trustworthy with production workloads.

Is anyone using open source models for anything major ?

Re: The Kimi K3 Moment

#159

Earlier quoted context omitted.

Is this some form of rage bait? 2005 we hadn't the GPUs, we have today. There are other factors, but I think this is the big one. The mathematics of building an LLM are really old, we just hadn't the hardware to do the needed calculations.

Right. Therefor it's not simply a derivative of information. The hardware is required to build the model. Software as well. The model uses information, it is not "distilled" from it. "Distillation" literally means to separate and take some components out of something. You can distill how a model works from a model. You cant distill a model from information because the information does not contain the model. People ar…

That argument is moot as distillation also requires a lot of hardware and software, if copying models was as easy as that, we would have hundreds of competing models.

Re: The Kimi K3 Moment

#160
post #8

I tried Kimi K3 on a task I've done with every other model I use regularly ( https://swelljoe.com/post/i-let-every-agent-implement-its-ow... ) and found it chewed a lot longer on the problem and ate up almost the entirety of a 5 hour usage limit on their $19 plan. I only have the $20 plan from OpenAI and the same task, with a lot of the same implementation details as Kimi Code, only took a few minutes and consumed al…

Earlier today I made Claude code implement a feature with fable. It worked roughly 60 minutes and used around 30% of my 100€ subs 5h sessions.

Then I typed /code-review in a second terminal/clean session after the analysis was done (no code changes) the usage was 99%. I then asked it to write that into a review.md so I could restart from that the next day. Sadly the last % wasn't enough for that.

Ymmv, these models behave very differently with no discernable reason. Usually reviews(even with fable) take like 10-20%... Yet suddenly you get it to burn through 65-69% in 15 minutes or so

Post reply on HN