Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
31–40 of 152 posts
Re: Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
#32or 40 macbooks with each 4 ssd. to get 40 tps.
Re: Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
#33A medium prompt in only 11 days.
Re: Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
#34A medium prompt in only 11 days.
Re: Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
#35Re: Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
#36But why though? Cannot possibly be useful at such slow speeds, and costs a ton to perform that badly
I've never understood why "Hacker" News so frequently gets "But why though?" comments at the top. The entire history of innovation is filled with people doing something just to see they can get it to work, even if badly, and then people continue to iterate on that until it works better, then works well, and then is so obvious people would never even question it. But it all starts with someone doing it to scratch an i…
What are you going to do with a computer? I've always hated this attitude. We do these things because they are interesting to us, for the fun of exploration, because we enjoy learning, because we want to iterate and improve, to make the world better, or any plethora of reasons that involve intellectual curiosity of some sort.
Re: Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
#37Re: Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
#38Reminds me of Deep Thought from Hitchhiker's Guide to the Galaxy
Re: Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
#39Re: Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
#40But why though? Cannot possibly be useful at such slow speeds, and costs a ton to perform that badly
Not useful for chat, agreed — and I wouldn't pretend otherwise. It's useful for the other kind of work: scheduled, unattended jobs where nobody is waiting on the cursor. My use is day/week/month end review — go through the numbers, flag what doesn't reconcile, draft the report — and there the two things that matter are that the model is good enough to trust with the judgement (K3 is, and it's the full 2.8T model, not…
Such of a Claudism. Not criticizing, just noticing.