0.01 tk/s is unusable for anything, you would wait a whole day for just 1000 token of output, what is the point of projects like this?
I like seeing the latest and greatest model crammed into new systems to see how it fares. To deal with the speed, one person on reddit suggested using it in an email interface rather than a chat interface.
Running Kimi K3 on a M1 Max
21–30 of 97 posts
Re: Running Kimi K3 on a M1 Max
#220.01 tk/s is unusable for anything, you would wait a whole day for just 1000 token of output, what is the point of projects like this?
Re: Running Kimi K3 on a M1 Max
#23I don't know if I'd call this "running"
Re: Running Kimi K3 on a M1 Max
#240.01 tk/s is unusable for anything, you would wait a whole day for just 1000 token of output, what is the point of projects like this?
(note. I am not a believer in AGI)
"useful" is highly contextual. The clock of the long "now" is not useful in the sense you mean, to synchronise your wristwatch. I'm still glad it exists.
Re: Running Kimi K3 on a M1 Max
#25> ~60–76 s/token I don't know if I'd call this "running"
UPD: I know it's not the same at all, just the reversal of units that gets me
Re: Running Kimi K3 on a M1 Max
#26> ~60–76 s/token I don't know if I'd call this "running"
Re: Running Kimi K3 on a M1 Max
#27Re: Running Kimi K3 on a M1 Max
#28> ~60–76 s/token I don't know if I'd call this "running"
I commented similarly below, but as a terrible programmer, I probably perform about 1 minute per token too (at Kimi 3 level). It puts into context how I think about intelligence
It’d be like thinking as slowly as Ents talk to each other in Lord of the Rings.
Re: Running Kimi K3 on a M1 Max
#29Re: Running Kimi K3 on a M1 Max
#30The title should probably be edited to specify "M1 Max" instead of "M1 Mac". You aren't running K3 on a base M1 anytime soon. Either way, still a very impressive project.