Live data from Hacker News

Kimi K2.6: Advancing open-source coding

kimi.com

361–370 of 394 posts

Re: Kimi K2.6: Advancing open-source coding

#361

Earlier quoted context omitted.

I think one of the motivations is undermining US companies. OpenAI and Anthropic are the two biggest players, and are American. Open weights models reduce the power those two big players have over the industry. If the Chinese companies tried to play by US rules and close-source their products then people would mostly use ChatGPT and Claude. So the Chinese companies don't make a ton of profit either way, but by releas…

Is Meta trying to keep the US from making as much profit with Llama? Is Google with Gemma? Microsoft with Phi? It's much simpler than some flag-waving nationalism.

Just because other companies have released open weights models doesn’t mean they are doing so with the same motivation.

And I never implied that the Chinese companies decision making was as simple as this. I said I think this is _one of_ the reasons.

Re: Kimi K2.6: Advancing open-source coding

#363

Earlier quoted context omitted.

At this point drawing these Pelicans must be in the training data sets.

Yes we all know that, but we still like to see the pelicans because it's a tradition more or less

Why no Utah Teapot!

Re: Kimi K2.6: Advancing open-source coding

#364

There is some humor in the fact that china (of all countries) is pioneering possibly the world's most important tech via open source, while we (US) are doing the exact opposite.

I think one of the motivations is undermining US companies. OpenAI and Anthropic are the two biggest players, and are American. Open weights models reduce the power those two big players have over the industry. If the Chinese companies tried to play by US rules and close-source their products then people would mostly use ChatGPT and Claude. So the Chinese companies don't make a ton of profit either way, but by releas…

It's a strategy so old it has a name: Commoditize your complement / competition

Also even a Joel Spolsky article (did he come up with the term?): https://www.joelonsoftware.com/2002/06/12/strategy-letter-v/

The Chinese want to kill a possible US monopoly in the crib. Yay for open source the old bane of monopolies.

Re: Kimi K2.6: Advancing open-source coding

#365

Earlier quoted context omitted.

I think one of the motivations is undermining US companies. OpenAI and Anthropic are the two biggest players, and are American. Open weights models reduce the power those two big players have over the industry. If the Chinese companies tried to play by US rules and close-source their products then people would mostly use ChatGPT and Claude. So the Chinese companies don't make a ton of profit either way, but by releas…

American companies just take those Chinese models and repackage them for profit like Cursors composer-2.

Smaller US companies that compete with the larger US companies, making monopoly in this market that much less likely.

Re: Kimi K2.6: Advancing open-source coding

#366

Earlier quoted context omitted.

> If a model finally comes out that produces an excellent SVG of a pelican riding a bicycle you can bet I’m going to test it on all manner of creatures riding all sorts of transportation devices. This relies on the false premise that, if they would include it in their training dataset, it would be perfect. All they need to do is be good enough and better than the other, not perfect.

I'm not sure if we can have a "perfect" Pelican riding a bicycle. Like, I could probably commission a highly experienced artist to draw one and I don't think it would be perfect. The legs would probably have to be too long, or pedals oddly placed, or handles strange, or wings with hands. Based on the one Simon commented though, I'd say we're in decent territory to try the latter part of his hypothesis.

> The legs would probably have to be too long, or pedals oddly placed, or handles strange, or wings with hands.

In all seriousness, that's what makes it an interesting test: it's asking for something technically impossible, that requires artistic license to make coherent.

Making specific choices on where to bend reality (and where not to) is a big chunk of visual art.

Re: Kimi K2.6: Advancing open-source coding

#367

There is some humor in the fact that china (of all countries) is pioneering possibly the world's most important tech via open source, while we (US) are doing the exact opposite.

We are at the point where uncontrolled capitalism collides with humanity. I do wonder where we go from here.

the chinese read marx and decided the only way is to overcome the limitations of capitalism through saturation of its potentialities under the rule of the workers party

Re: Kimi K2.6: Advancing open-source coding

#369

Earlier quoted context omitted.

I'm genuinely so grateful for them $200/m minimum to use Claude would bankrupt my country's white collar labor market

I would really appreciate a response because I'm sure you know that Anthropic has at least two lower priced tiers before the $200/m one, so I assume the $200/m tier is necessary because you use it heavily? Now given that the $200/m Tier is the most heavily (I believe at 20x?) subsidized tier, How or what are you using instead that achieves comparable good enough performance for a fraction of the price? I've heard GLM…

I’m currently on the $100/m plan and my usage limits get exhausted every week even though I’m not using it for full time work

I can’t imagine how little mileage you get out of the $20/month plan

For context, $250/month is the starting salary of an engineering hire at my country’s biggest IT company. Even $100/m is beyond the ability of any student or early professional to pay out of pocket

Re: Kimi K2.6: Advancing open-source coding

#370

Early benchmarks show tremendous improvement over Kimi K2 Thinking, which didn't perform well on our benchmarks (and we do use best available quantization). Kimi K2.6 is currently the top open weights model in one-shot coding reasoning, a little better than GLM 5.1, and still a strong contender against SOTA models from ~3 months ago (comparable to Gemini 3.1 Pro Preview). Agentic tests are still running, check back t…

Cool website. I don't understand enough about the various benchmarks or how they're done to judge whether or not anything is accurate, but I love the layout and features especially the spectator feature which is pretty cool. One thing, I saw the "Market simulator" spectator feature but didn't see a corresponding benchmark for that. Is it "Finance" or "Betting" or "Trading"?
Post reply on HN