New Mac Studio with M5 Max and M5 Ultra
141–150 of 571 posts
Re: New Mac Studio with M5 Max and M5 Ultra
#142Sorry for being lazy, but is there a rough breakdown like "You get sonnet level for M5 and Opus for M5 pro, etc.", or is it still speculative. Or put simpler, do you get Opus level for the 256GB M5 Max?
Re: New Mac Studio with M5 Max and M5 Ultra
#143Earlier quoted context omitted.
The relevant comparison isn't one mac studio to one RTX 6000, it's a 24 channel DDR5 system, which also has ~1.2TB/s of memory bandwidth (or more when Xeon 6 compatible 8800mt/s memory becomes widely available), vastly higher prefill due to more CPU horsepower, orders of magnitude faster networking, can hook into GPU accelerators, can be upgraded etc. A baseline 384GB system from eg Puget is ~30K vs ~12K for the 256G…
So 3x more, plus the cost of a GPU (another 10k?). How is that value for money to get slightly better performance?
If you just want to run Qwen 3.8 27B and Deepseek v4 Flash in perpetuity and that's it, there are a lot of solutions that will work and this is a fairly user friendly one.
Re: New Mac Studio with M5 Max and M5 Ultra
#144Mac Studio with M5 Max starts at $2,499 (U.S.) and $2,299 (U.S.) for education. Additional configure-to-order options are available at apple.com/mac-studio. Mac Studio with M5 Ultra starts at $5,499 (U.S.) and $5,099 (U.S.). I am in Europe, and the Mac Studio M5 Ultra GPU 64 cores with 96GB RAM is up to 6.649,00 €. Ouch.
Don't forget that US prices usually do not include the VAT, while EU prices usually do include respective VAT.
Re: New Mac Studio with M5 Max and M5 Ultra
#145Earlier quoted context omitted.
I've been thinking about this a fair bit recently. We make a lot of price/performance compromises for having an attached screen and keyboard on our computer. That was what got me started. Then I remembered the days of having to go to a special corner of the house to use a computer, vs now when I have a computer with me all the time. In my bag, on the sofa, on the train. Hell, I'm writing this on the work MBP while wa…
I've done that, but I chose a Framework Desktop instead. The latest Fedora is closer to Snow Leopard than anything Apple has to offer. Downside is that I still have my MBP because of the lock-in and occasionally pick it up to do computer stuff in weird places.
If your computing needs line up, it's a very serviceable approach.
Re: New Mac Studio with M5 Max and M5 Ultra
#146Re: New Mac Studio with M5 Max and M5 Ultra
#147Sorry for being lazy, but is there a rough breakdown like "You get sonnet level for M5 and Opus for M5 pro, etc.", or is it still speculative. Or put simpler, do you get Opus level for the 256GB M5 Max?
Re: New Mac Studio with M5 Max and M5 Ultra
#148Mac Studio with M5 Max starts at $2,499 (U.S.) and $2,299 (U.S.) for education. Additional configure-to-order options are available at apple.com/mac-studio. Mac Studio with M5 Ultra starts at $5,499 (U.S.) and $5,099 (U.S.). I am in Europe, and the Mac Studio M5 Ultra GPU 64 cores with 96GB RAM is up to 6.649,00 €. Ouch.
> I am in Europe, and the Mac Studio M5 Ultra GPU 64 cores with 96GB RAM is up to 6.649,00 €. Ouch. It would be significantly cheaper to fly to a tariff-free country and buy there.
Re: New Mac Studio with M5 Max and M5 Ultra
#149Earlier quoted context omitted.
I've been holding out, because I think my next purchase will be a Studio with an Ultra Chip in it. I'm wanting it to be a "forever" server, so I'm holding out while I can.
Supply rumours are next year we see an M7 AI-focused chip with large inference performance upgrades. It's unlikely we'll see heavy upgrades in other areas. If you care about AI, it's worth waiting. If you don't, pull the trigger now. RAM constraints are likely to get worse next year. Or wait 2-3 years and prices should be back to Earth (plus newer and even better chips). I'm waiting this out.
Re: New Mac Studio with M5 Max and M5 Ultra
#150For a non-quantized Deepseek V4 flash on an ultra, I would estimate about 1000+ tokens per second prefill and 50+ tokens per second on generation. This is actually quite usable and near parity to cloud.
They mention "adds the GPU Neural Accelerators." which, if exploitable for LLM loads, would probably help the prefill a lot