Earlier quoted context omitted.
Whether or not it's propaganda is different from the fact that it is owned by the CCP.
Doesn't matter, because they're open-weight, so I can just download them to my PC and... hey, look, now they're owned by me! Unlike the "good" Western counterparts which are all fully proprietary. (Except Mistral, but they're nowhere near SOTA.)
Kimi K2.7-Code: open-source coding model with better token efficiency
141–150 of 254 posts
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#142Has anyone taken these open weight models from China and stripped the CCP out of them? I do not mean that snarkily, I mean review them thoroughly using techniques for weight introspection (concept activations) in response to things that one might expect would trigger deceptive/malicious behavior if the CCP had actually tried to implant context-specific behaviors (e.g. the accusation of generating vulnerable code if b…
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#143Has anyone taken these open weight models from China and stripped the CCP out of them? I do not mean that snarkily, I mean review them thoroughly using techniques for weight introspection (concept activations) in response to things that one might expect would trigger deceptive/malicious behavior if the CCP had actually tried to implant context-specific behaviors (e.g. the accusation of generating vulnerable code if b…
The CCP is not influencing my Rust code quality that much. Though I did notice all my lifetimes are now 'static because nothing is ever allowed to leave the party's ownership, unsafe blocks require approval from a central committee.
Honestly the scariest part is that shared mutable state is forbidden unless the state is doing the sharing.
Otherwise it is pretty ok.
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#144Has anyone taken these open weight models from China and stripped the CCP out of them? I do not mean that snarkily, I mean review them thoroughly using techniques for weight introspection (concept activations) in response to things that one might expect would trigger deceptive/malicious behavior if the CCP had actually tried to implant context-specific behaviors (e.g. the accusation of generating vulnerable code if b…
Eh even corporate created LLMs are suspect to corporate biases. Nothing is safe.
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#145Earlier quoted context omitted.
I really hope we stop using the term "Chinese models". It has this air of Negative connotation. It's the equivalent of calling cars Japanese, which people used to do but now is almost entirely meaningless. You just call them Toyota, Honda, Lexus etc.
[flagged]
[1] https://www.tomshardware.com/tech-industry/artificial-intell...
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#146Personally, when I use open code or routers, I feel that beyond a certain level, the models don't make a huge difference to me. Except for expensive and mediocre models like Gemini. In that sense, Chinese models are pretty good. I usually write code in function or method units and then design and assemble them together. GPT series models are more thorough and better, but I'm not sure if the difference is enormous. It…
I've kind of given up on the routers for "free" inference, as you would expect, they tend to give you sub-par thinking because they are obviously trying to conserve as much inference as possible. I've had some success turning my macbook M1 pro into a heating pad with Qwen 3.6 35B A3B MTP. Trying to use Gemini models "locally" resulted in a similar "short shrift" of effort resulting in mistakes and lots of turns. The…
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#147In OpenRouter, there is an "int4" tag for Moonshot provider of Kimi K2. 7 Code. Isn't that too low, particularly coming from the very developer of the model? Os that a mistake? How is it in their direct API offer?
The model is natively quantized (i.e. it was trained that way in the first place, so this is not a post-training quantization which degrades performance).
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#148Earlier quoted context omitted.
yes, yes, the spectre of communism, BYD is the CCP, Alibaba is the CCP, stealing your children and eating them for Mao, bla bla bla. I have a feeling you'd be slightly salty at people saying "Google and Tesla are making CIA models"
Google and Tesla making products to sell to the government is different than the government funding the government to make products for the government. In China it's all one entity with these mock facades of privatization. Trump cannot instruct Google to put picture of dogs on their homepage. If Xi wakes up and wants dogs on Alibaba's homepage, give it 30 minutes. It's wholly ignorant or dishonest to make the compari…
Tim Apple and the other tech CEO constantly groveling at Trump’s feet indicates that he might be able to do that.
Just like threatening TV networks about having their licenses revoked of blocking mergers unless they fire the people making fun of him on TV (of course with slightly mixed success)
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#149Earlier quoted context omitted.
> Trump cannot instruct Google to put picture of dogs on their homepage. Sundar Pichai would personally be barking on a livestream on the homepage. Trump is quite literally the one president showing that the US has zero rules or anything to hold power back from the white house, really not the example you want.
Seems like everyday Trump has another order struck down by the courts. Sundar can do whatever he wants, but he has no legal obligation to do any of it.
e.g. he had Colbert fired (and who knows what else) by threatening to block the Paramount/Skydance merger
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#150Earlier quoted context omitted.
This is the cursor callout. Don't make us shame you into disclosure
Shaming others when all AI is trained off scraped content and code huh? Many of those sources either breaking ToS or being illegal, such as Anna’s Archive. Bold move. And Chinese models in particular have been accused of distilling off American models. Don’t you know there’s no honor among thieves?