0.01 tk/s is unusable for anything, you would wait a whole day for just 1000 token of output, what is the point of projects like this?
16 tokens / s is not nothing.
Running Kimi K3 on a M1 Max
31–40 of 97 posts
Re: Running Kimi K3 on a M1 Max
#320.01 tk/s is unusable for anything, you would wait a whole day for just 1000 token of output, what is the point of projects like this?
So you subscribe to the belief we won't in future find mentalism in other galaxies or solar systems which operate on mechanisms we don't understand and think v e r y s l o w w w w w w w l y ? (note. I am not a believer in AGI) "useful" is highly contextual. The clock of the long "now" is not useful in the sense you mean, to synchronise your wristwatch. I'm still glad it exists.
Re: Running Kimi K3 on a M1 Max
#33Anyone who knows the state of NVMe hardware more than me know if this would obliterate the lifespan of your drive? Seems like the biggest limitation to me (some people are probably fine with letting their Macs churn over the weekend).
Re: Running Kimi K3 on a M1 Max
#34Re: Running Kimi K3 on a M1 Max
#35Earlier quoted context omitted.
I commented similarly below, but as a terrible programmer, I probably perform about 1 minute per token too (at Kimi 3 level). It puts into context how I think about intelligence
Maybe in terms of code produced, but one token is only a fragment of a thought for an LLM. It’d be like thinking as slowly as Ents talk to each other in Lord of the Rings.
Re: Running Kimi K3 on a M1 Max
#36Re: Running Kimi K3 on a M1 Max
#37idk how people access (soldout) and even afford 512GB RAM MacStudio's. Isn't it $40k or so?
Re: Running Kimi K3 on a M1 Max
#38SSD streaming on an M5 Max 128GB: https://x.com/antirez/status/2082136334160818528 Soon decent speed across two Mac Studios with 512GB of RAM.
Re: Running Kimi K3 on a M1 Max
#39Exactly my machine 64GB M1 Max So happy about this! ♡ idk how people access (soldout) and even afford 512GB RAM MacStudio's. Isn't it $40k or so?