Live data from Hacker News

Kimi K2.6: Advancing open-source coding

kimi.com

141–150 of 394 posts

Re: Kimi K2.6: Advancing open-source coding

#141
post #129

There is some humor in the fact that china (of all countries) is pioneering possibly the world's most important tech via open source, while we (US) are doing the exact opposite.

I wonder if there's a strategy behind all of this on China's side. I know the CCP uses a direct hand in many affairs in China, but is there an actual coordinated effort to compete with, or sabotage the West?

Chinese AI companies want investors too. Nobody would believe they can compete with western companies unless they release something you can run on your own hardware.

After all historically both statistics and research that comes out of China is not very trustworthy.

Re: Kimi K2.6: Advancing open-source coding

#142

Earlier quoted context omitted.

The psyop continues. Mythos until it’s released is vaporware. Notice how you can try kimi 2.6. Where is the same for mythos?

At this point it seems more like the result of a psyop to presume that a new anthropic model should be considered vaporware until released.

[deleted]

Re: Kimi K2.6: Advancing open-source coding

#143

Earlier quoted context omitted.

Only if you use Kimi API directly - the censorship is done externally. The model itself talks fine about Tiananmen, you can check on Openrouter. There might be less visible biases, though.

That's what I wrote? Except that it also clearly has internal bias?

Everything has some sort of bias. Most text is written by those who like writing.

Re: Kimi K2.6: Advancing open-source coding

#144

Earlier quoted context omitted.

Only if you use Kimi API directly - the censorship is done externally. The model itself talks fine about Tiananmen, you can check on Openrouter. There might be less visible biases, though.

That's what I wrote? Except that it also clearly has internal bias?

> That's what I wrote?

No.

You wrote that "you won't hear about Tiananmen square from this model" and atemerev wrote that "the model itself talks fine about Tiananmen".

You wrote that "it can easily access any withheld or missing info from training data via tool calls" and atemerev wrote that "the model itself talks fine about Tiananmen".

Re: Kimi K2.6: Advancing open-source coding

#145
post #83

Are there any coding plans for this? (aka no token limit, just api call limit). Recently my account failed to be billed for GLM on z.ai and my subscription expired because of this... the pricing for GLM went through the roof in recent months, though...

Kimi has their own subscription that works basically the same as all the others.

https://www.kimi.com/code

Re: Kimi K2.6: Advancing open-source coding

#146

Earlier quoted context omitted.

Why use China model API from China if there are many independent providers available via Openrouter?

Openrouter will route to china hosted models when there are US hosted providers of the same model. Is there a setting to set your preference or to blacklist providers like alibaba cloud for example? I use OpenCode and the openrouter provider. From opencode I only select the model like kimi-2.6 and have no way of selecting which cloud hosting will receive my request.

Settings > Guardrails > [your workspace] > Providers + Block provider

Re: Kimi K2.6: Advancing open-source coding

#147

Earlier quoted context omitted.

Why use China model API from China if there are many independent providers available via Openrouter?

Openrouter will route to china hosted models when there are US hosted providers of the same model. Is there a setting to set your preference or to blacklist providers like alibaba cloud for example? I use OpenCode and the openrouter provider. From opencode I only select the model like kimi-2.6 and have no way of selecting which cloud hosting will receive my request.

Yes, you can globally ban providers in your openrouter settings.

Re: Kimi K2.6: Advancing open-source coding

#148

Earlier quoted context omitted.

Why use China model API from China if there are many independent providers available via Openrouter?

Openrouter will route to china hosted models when there are US hosted providers of the same model. Is there a setting to set your preference or to blacklist providers like alibaba cloud for example? I use OpenCode and the openrouter provider. From opencode I only select the model like kimi-2.6 and have no way of selecting which cloud hosting will receive my request.

Yes, you can blacklist providers in OpenRouter account settings.

Re: Kimi K2.6: Advancing open-source coding

#149
post #31

Earlier quoted context omitted.

This should erase any doubt that AI Labs are making $$$ on API inference. Kimi 2.5 (which this is based on) is served at $0.44 input / $2 output by a ton of different providers on OpenRouter, 2.6 will certainly be similar. That's about 11X less than Opus for similar smarts.

How does it erase any doubt? You’re implying Chinese things can’t be actually cheaper to produce than American which is laughable

Most of those inference providers are American, and China is actually at a disadvantage here because of export restrictions - US companies are using newer and more efficient chips.

Re: Kimi K2.6: Advancing open-source coding

#150

Beats Opus and Open Source? I really hope this holds true in real world use cases as well and not only benchmarks. Congrats to Kimi team!

K2.6-code-preview was a minor, but noticeable jump, especially in a long running testing task and prior Moonshot releases have been the only models that I'd consider a suitably competitive replacement for Anthropic models. The way they approach tool calls, task inference and adherence is far closer than any other providers output, similar to how GLM models map far more closely to OpenAIs releases. Whether task adherence, task assessment, task evaluation or task inference, K2.5 got closer to Opus 4.5 than any other model (but was still behind overall).

I will have to test this full release of K2.6 but could see it serve as a very good overall drop-in replacement for Opus 4.5 and Opus 4.6 at 200k across the vast majority of tasks.

I will say however that Opus 4.7 Max 1M has been a very significant jump in performance for me, especially in tasks beyond 120k token where I'd argue it is now the most reliable model in continued task adherence and tool calling without compaction. Ironically, my initial experience was less than pleasant as on XHigh I found task adherence to have regressed even with less than 1/10th of the context window having been used.

Am very interested in K2.6s compaction strategy (which appears to be very simply all things considered) and how it performs beyond 100k tokens. As it stands, only OpenAI models have made compaction for long running tasks work well, though overall, GPT-5.4 is still inferior in my tests regardless of context window over other models such as Opus 4.6 1m and Opus 4.7 1m. Haven't gotten around to testing Opus 4.7 200k and will have to do this to properly assess K2.6 fairly, but I'd be very surprised if K2.6 truly beat Opus 4.7 200k given the jump I have experienced.

Post reply on HN