Live data from Hacker News

Nvidia DGX Spark and Apple Mac Studio = 4x Faster LLM Inference with EXO 1.0

blog.exolabs.net

21–22 of 22 posts

Re: Nvidia DGX Spark and Apple Mac Studio = 4x Faster LLM Inference with EXO 1.0

#21

It’s really sad that exo went private.

How do you know this happened? I thought it was an abandoned project until I saw this post. I've been diligently checking weekly for new releases but nothing for almost a year...

Appreciate you checking back so often. We have some exciting plans. Keep checking and it won't be long before something pops up :)

Re: Nvidia DGX Spark and Apple Mac Studio = 4x Faster LLM Inference with EXO 1.0

#22
post #14

Wouldn't this restrict memory to 128GB, wasting M3 Ultra potential?

Blog author here. Actually, no. The model can be streamed into the DGX Spark, so we can run prefill of models much larger than 128GB (e.g. DeepSeek R1) on the DGX Spark. This feature is coming to EXO 1.0 which will be open-sourced soonTM.

Excellent! Good luck!
Post reply on HN