Live data from Hacker News

Kimi K2 1T model runs on 2 512GB M3 Ultras

twitter.com

121–125 of 125 posts

Re: Kimi K2 1T model runs on 2 512GB M3 Ultras

#121
post #38

Earlier quoted context omitted.

? The user directly addresses this.

K2 and K2T are drastically different models released a significant amount of time apart, with wildly different capabilities and post training. K2T is much closer in capability to 4.5 Sonnet from what I've heard.

[deleted]

Re: Kimi K2 1T model runs on 2 512GB M3 Ultras

#122
post #96

Earlier quoted context omitted.

All of those things you can still do renting AI server compute though? I think privacy and cool-factor are the only real reasons why it would be rational for someone to spend checks the apple store $19,000 on computer hardware...

Why do you look at this as a consumer? Have you never heard of businesses spending money on hardware???

And what reasons would a business have to spend the money on hardware instead of cloud services? Privacy

Re: Kimi K2 1T model runs on 2 512GB M3 Ultras

#123
post #88
post #81

Earlier quoted context omitted.

Mistral Large 3 is reportedly using Deepseek V3.2 architecture with larger experts and fewer of them, and a 2B params vision module.

According to whom? I haven't seen any claims of that being the case (other than you), just that there are similar decisions made by both of them. https://mistral.ai/news/mistral-3

https://www.reddit.com/r/LocalLLaMA/comments/1plpc6h/mistral...

Re: Kimi K2 1T model runs on 2 512GB M3 Ultras

#124

Earlier quoted context omitted.

> What does that leave us with? At the start, with no benchmark. Because LLMs can't reason at this time, and because we don't have a reliable way of grading LLM reasoning, and because people are stubborn thinking LLMs are actually reasoning we're at the start. When you ask a LLM "2 + 2 = ", it doesn't add the numbers together, it just looks up one of the stories it memorized and return what happens next. Probably in…

That's kind of nonsense, since if I ask you what's five times six, you don't do the math in your head, you spit out the value of the multiplication table you memorized in primary school. Doing the math on paper is tool use, which models can easily do too if you give them the option, writing adhoc python scripts to run the math you ask them to with exact results. There is definitely a lot of generalization going on be…

On the other hand, if you ask me what five times six in base eight is, I can spend a second and repy thirtysix. Is there an LLM able to do that yet?

Re: Kimi K2 1T model runs on 2 512GB M3 Ultras

#125
post #122

Earlier quoted context omitted.

Why do you look at this as a consumer? Have you never heard of businesses spending money on hardware???

And what reasons would a business have to spend the money on hardware instead of cloud services? Privacy

Seriously?? You’ve never seen a company want to control its entire stack and hardware for ANY reason but privacy? Cloud is great, but it doesn’t fit every use case.
Post reply on HN