Due to pricing insanity (not that Apple prices weren’t insane before the ram/ssd shortages) I’m not in the market but I do wonder if my next computer should be a Mac Studio instead of a MBP that lives its life docked. Might be better to just run a Studio and Neo for the very few times I actually need remote capabilities.
I’ve wanted the idea of a home server or even home “mainframe” and some terminals for me and family forever but the idea never takes off….
New Mac Studio with M5 Max and M5 Ultra
331–340 of 571 posts
Re: New Mac Studio with M5 Max and M5 Ultra
#332> Apple’s most powerful Mac raises the bar for local AI Wow, "Local AI" mentioned in the subheading above the fold - it's really awesome to see Apple leaning into this use case and I think it will definitely pay off for them going forward. Fingers crossed Apple is able to put some engineering effort towards shipping with one of the frontier open weight models included and optimized exactly for the machine.
I actually wouldn't want a serious model shipping 'on disk'. Models release so often, it's going to get outdated quickly. LMStudio is trivial to set up.
Re: New Mac Studio with M5 Max and M5 Ultra
#333Due to pricing insanity (not that Apple prices weren’t insane before the ram/ssd shortages) I’m not in the market but I do wonder if my next computer should be a Mac Studio instead of a MBP that lives its life docked. Might be better to just run a Studio and Neo for the very few times I actually need remote capabilities.
I've been holding out, because I think my next purchase will be a Studio with an Ultra Chip in it. I'm wanting it to be a "forever" server, so I'm holding out while I can.
Re: New Mac Studio with M5 Max and M5 Ultra
#334Sorry for being lazy, but is there a rough breakdown like "You get sonnet level for M5 and Opus for M5 pro, etc.", or is it still speculative. Or put simpler, do you get Opus level for the 256GB M5 Max?
For local LLMs with a Mac, rule of thumb is you always want an Ultra (due to memory bandwidth). Even an M1 Ultra is superior to an M6 Pro in this regard. There are no configurations even close to running something comparable to frontier model variants, they're simply far too large, but something like full precision Qwen 35b or DeepSeek 70b at 50+ t/s is well within available configuration, and potential for plenty of…
I'm using Flash heavily, and I would describe it as nearly as intelligent as Sonnet-class in agentic coding, but more usable. Less world knowledge of course, and definitely a bit less intelligent; but not _that_ much.
On usability: Takes less handholding, less likely to make unsolicited refactors or whatever, and the writing style is readable.
It's not great at super-long-horizon goals as the Claude 5 models are; but if you have a good harness, you can get around that.
Re: New Mac Studio with M5 Max and M5 Ultra
#335Re: New Mac Studio with M5 Max and M5 Ultra
#3361.2 TB/s bandwidth of M5 Ultra comes from two dies of M5 Max (each 614 GB/s) connected together using 4.4 TB/s inter-die fabric. For a non-quantized Deepseek V4 flash on an ultra, I would estimate about 1000+ tokens per second prefill and 50+ tokens per second on generation. This is actually quite usable and near parity to cloud. They mention "adds the GPU Neural Accelerators." which, if exploitable for LLM loads, wo…
How much would is the cost for that machine though, I'm pretty sure I could just buy tokens from a provider and never run out of money for 10 years, and get far better quality output because inference is being served by professionals on far better hardware and this machine would be obsolete long before that as well. Hosting local seems like a possibly the dumbest thing you could possibly do from an economics perspect…
Re: New Mac Studio with M5 Max and M5 Ultra
#337Re: New Mac Studio with M5 Max and M5 Ultra
#338Earlier quoted context omitted.
I would not use Neo, Air with 24 GB ram should be your minimum because you will want to do some work locally even when you're running a VPN/SSH thin client setup. I'm currently SSH-ing into my workstation from my M4 Air with 24 gb ram and it's ideal for this flow. Slack/editors/clients/browsers/etc. easily gobble up over 16 GB. I have no more dev tools/compilers/source on my client machines, everything is dockered up…
Can that run qwen? Also can you explain what you meant by the downside ?
I meant the only downside of not using a pro for my use-case is that I don't have a HDMI port on device with high refresh rate. I have a decent dock that can give me 4k/60hz HDMI but not 120hz refresh for macs (it works on windows). It's a minor thing but having 120hz is nice when I'm docked at home.
Re: New Mac Studio with M5 Max and M5 Ultra
#339Due to pricing insanity (not that Apple prices weren’t insane before the ram/ssd shortages) I’m not in the market but I do wonder if my next computer should be a Mac Studio instead of a MBP that lives its life docked. Might be better to just run a Studio and Neo for the very few times I actually need remote capabilities.
I've been thinking about this a fair bit recently. We make a lot of price/performance compromises for having an attached screen and keyboard on our computer. That was what got me started. Then I remembered the days of having to go to a special corner of the house to use a computer, vs now when I have a computer with me all the time. In my bag, on the sofa, on the train. Hell, I'm writing this on the work MBP while wa…
Re: New Mac Studio with M5 Max and M5 Ultra
#340Due to pricing insanity (not that Apple prices weren’t insane before the ram/ssd shortages) I’m not in the market but I do wonder if my next computer should be a Mac Studio instead of a MBP that lives its life docked. Might be better to just run a Studio and Neo for the very few times I actually need remote capabilities.
Thus, my current workhorse is a Studio.