Live data from Hacker News

Apple M3 Ultra

apple.com

751–760 of 1001 posts

Re: Apple M3 Ultra

#751
post #349

Earlier quoted context omitted.

The M4 Pro is 56% faster in ST performance against AMD’s new Strix Halo while being 3.6x more efficient. Source: https://www.notebookcheck.net/AMD-Ryzen-AI-Max-395-Analysis-... Cinebench 2024 results.

That’s a laptop part, so it makes different tradeoffs. Somewhere on the internet there is a tdp wattage vs performance x-y plot. There’s a pareto optimal region where all the apple and amd parts live. Apple owns low tdp, AMD owns high tdp. They duke it out in the middle. Intel is nowhere close to the line. I’d guess someone has made one that includes datacenter ARM, but I’ve never seen it.

[deleted]

Re: Apple M3 Ultra

#752

Earlier quoted context omitted.

Not to get in the way of good snark or anything. But.. Apple isn't _requiring_ that everyone uses MacOS on their systems. But you have to bring your own engineering effort to actually make another OS run. And so far Asahi is the only effort that I'm aware of (there were alternatives in the very beginning, but they didn't even get to M2 right?)

> But you have to bring your own engineering effort to actually make another OS run. I mean, that's usually how it works though. When IBM launched the PS/2, they didn't support anything other than PC-DOS and OS/2, Microsoft had to make MS-DOS work for it (I mean... they did get support from IBM, but not really), the 386BSD and Linux communities brought the engineering effort without IBM's involvement. When Apple was…

When Apple was making Motorola Macs, they may have given Be a little help, but didn't support any other OSes that appeared. Same with PowerPC.

Apple briefly supported a Linux distribution on PowerPC Macs: https://en.wikipedia.org/wiki/MkLinux.

Re: Apple M3 Ultra

#753

Thunderbolt 5 (TB 5) is pretty handy, you can have a very thin and lightweight laptop, then can get access to external GPU or eGPU via TB 5 if needed [1]. Now you can have your cake (lightweight laptop) and eat it too (potent GPU). [1] Asus just announced the world’s first Thunderbolt 5 eGPU: https://www.theverge.com/24336135/asus-thunderbolt-5-externa...

Apple Silicon does not work with eGPU.

Re: Apple M3 Ultra

#754

Computers these days - the more appealing, exciting, cooler desirable, the higher the price, into the stratosphere. $9499 What ever happening to competition in computing? Computing hardware competition used to be cut throat, drop dead, knife fight, last man standing brutally competitive. Now it's just a massive gold rush cash grab.

You take the top price of the top of the line newest pro chip apple produces and then make this argument?

Re: Apple M3 Ultra

#755

Earlier quoted context omitted.

A100 has 10x or so higher mem bandwidth

Per nvidia [1] A100 has memory bandwidth up to 2,039. So not 10x, more like 2x. [1] https://www.nvidia.com/content/dam/en-zz/Solutions/Data-Cent...

That's for 1, they were asking about 8x A100s, so 16x. H100 is double again.

Re: Apple M3 Ultra

#757
post #560

Earlier quoted context omitted.

They didn't increase the memory bandwidth. You can get the same memory bandwidth, which is available on the M2 Studio. Yes, yes, of course you can get 512 gigabytes of uRAM for 10 grand. The the question is if a llm will run with usable performance at that scale? The point is there's diminishing returns despite having enough uRAM with the same amount of memory bandwidth even with increased processing speed of the new…

Since no one specifically answered your question yet, yes, you should be able to get usable performance. A Q4_K_M GGUF of DeepSeek-R1 is 404GB. This is a 671B MoE that "only" has 37B activations per pass. You'd probably expect in the ballpark of 20-30 tok/s (depends on how much actually MBW can be utilized) for text generation. From my napkin math, the M3 Ultra TFLOPs is still relatively low (around 43 FP16 TFLOPs?),…

[deleted]

Re: Apple M3 Ultra

#758
post #734

Earlier quoted context omitted.

GH200 is nowhere near $343,000 number. You can get a single server order around 45k (with inception discount). If you are buying bulk, it goes down to sub-30k ish. This comes with a H100's performance and insane amount of high bandwith memory.

They probably meant 8xH200 for $343,000 which is in the ballpark.

Yes this is what I meant since 8 would cover 512GB of Ram

Re: Apple M3 Ultra

#759

Earlier quoted context omitted.

They didn't increase the memory bandwidth. You can get the same memory bandwidth, which is available on the M2 Studio. Yes, yes, of course you can get 512 gigabytes of uRAM for 10 grand. The the question is if a llm will run with usable performance at that scale? The point is there's diminishing returns despite having enough uRAM with the same amount of memory bandwidth even with increased processing speed of the new…

Any idea what the sRAM to uRAM ratio is on these new GPUs ? If they have meaningfully higher sRAM than the Hopper GPUs, it could lead to meaningful speedups in large model training. If they didn't increase the memory bandwidth, then 512GB will enable longer context lengths and that's about it right? No speedups For any speedups You may need some new variant of FlashAttention3 or something along similar lines to be pu…

I don't know what you mean by s and u, but there is only one kind of memory in the machine, that's what unified memory means.
Post reply on HN