Live data from Hacker News

Apple introduces M6 and M5 Ultra

apple.com

21–30 of 1001 posts

Re: Apple introduces M6 and M5 Ultra

#21
post #7
post #2

> M5 Ultra features a massive amount of high-bandwidth unified memory, up to 512GB, and delivers a staggering 1.2TB/s of unified memory bandwidth that is 50 percent higher than M3 Ultra. Apple never needed to participate in the AI race to zero. Because they were already at the finish line years ago building their own chips that can run large >100B parameter AI models locally.

1.2TB/s is 2/3 the speed of an nVidia 5090. But you get a generic computer and much more RAM. And you lose a couple of organs.

The real downside for me is not having Linux support.

It would take Apple one or two engineers to make Linux life much easier on macs. But Linux is outside their walled garden so it's ignored.

Re: Apple introduces M6 and M5 Ultra

#22
post #2

> M5 Ultra features a massive amount of high-bandwidth unified memory, up to 512GB, and delivers a staggering 1.2TB/s of unified memory bandwidth that is 50 percent higher than M3 Ultra. Apple never needed to participate in the AI race to zero. Because they were already at the finish line years ago building their own chips that can run large >100B parameter AI models locally.

I was so blown away at all the discourse surrounding "Apple fumbling on models". They should never have been in the model game to begin with. Apple crushes hardware over the last decade and that's a huge advantage today. In the end, massive models have proven to be very strong, but small models have proven to be good enough (especially with the recent Qwen 2.8 27B drop) and that's where I imagine the future will lie for consumers.

Re: Apple introduces M6 and M5 Ultra

#23
post #4
post #2

> M5 Ultra features a massive amount of high-bandwidth unified memory, up to 512GB, and delivers a staggering 1.2TB/s of unified memory bandwidth that is 50 percent higher than M3 Ultra. Apple never needed to participate in the AI race to zero. Because they were already at the finish line years ago building their own chips that can run large >100B parameter AI models locally.

Is there anything comparable that runs Linux, doesn't necessarily look as good, but is perhaps (a lot) cheaper/fixable? Or is this really pretty optimal? I mean this is not nvidia based right? It's all custom? So we can use it under Asahi perhaps? I want to get something for my company to run local models, wondering what would be a good option.

Strix platform maybe?

Re: Apple introduces M6 and M5 Ultra

#26
post #5

I know they are a phone company, but I think they should focus on local models software, not only hardware.

Software wise there's plenty to choose from already. Ollama/llama.cpp, LM Studio, Lemonade, vllm etc. Anything Apple would bring to the table?

Re: Apple introduces M6 and M5 Ultra

#28
post #5

I know they are a phone company, but I think they should focus on local models software, not only hardware.

Apple has the mlx framework. Most/all major software for running models locally support it. Apple also has RDMA for interconnecting multiple machines across Thunderbolt connections.

Re: Apple introduces M6 and M5 Ultra

#29
post #5

I know they are a phone company, but I think they should focus on local models software, not only hardware.

Uhhh, they are? They’re a hardware co for sure, but to say they aren’t focusing on on-device models is absurd on its face. They’ve spent over 2 years on Siri AI which is (mostly) local.

Re: Apple introduces M6 and M5 Ultra

#30
post #4
post #2

> M5 Ultra features a massive amount of high-bandwidth unified memory, up to 512GB, and delivers a staggering 1.2TB/s of unified memory bandwidth that is 50 percent higher than M3 Ultra. Apple never needed to participate in the AI race to zero. Because they were already at the finish line years ago building their own chips that can run large >100B parameter AI models locally.

Is there anything comparable that runs Linux, doesn't necessarily look as good, but is perhaps (a lot) cheaper/fixable? Or is this really pretty optimal? I mean this is not nvidia based right? It's all custom? So we can use it under Asahi perhaps? I want to get something for my company to run local models, wondering what would be a good option.

AFAIK, apple does not release drivers open source, asahi is a reverse-engineering endeavour and does not support GPU. For nvidia, there are both proprietary and open-source linux drivers. CUDA and inference works on linux with nvidia. I would recommend checking out this video of Alex Ziskind to shop for a computer to run local LLMs: https://www.youtube.com/watch?v=mevUEQcumzU&t=224s. TL;DR besides Apple he recommends, DGX Spark, Tenstorrent Wormhole N300, AMD Radeon 7900 and NVIDIA RTX 5090.
Post reply on HN