Live data from Hacker News

Nvidia DGX Spark

nvidia.com

91–100 of 222 posts

Re: Nvidia DGX Spark

#92
post #24

The mainstream options seem to be Ryzen AI Max 395+, ~120 tops (fp8?), 128GB RAM, $1999 Nvidia DGX Spark, ~1000 tops fp4, 128GB RAM, $3999 Mac Studio max spec, ~120 tflops (fp16?), 512GB RAM, 3x bandwidth, $9499 DGX Spark appears to potentially offer the most token per second, but less useful/value as everyday pc.

Maybe the real value of the DGX spark is to work on Switch 2 emulation. ARM + Nvidia GPU. Start with Switch 2 emulation on this machine and then optimize for others. (Yeah, I know, kind of expensive toy).

Re: Nvidia DGX Spark

#93

Paper launch. The people I know there who I have asked about it haven't even seen one yet

Ordered one in spring. Delivery time was pushed from July to September. Apparently they had a bug in the HDMI output.

Re: Nvidia DGX Spark

#94
post #67

Earlier quoted context omitted.

This thing has a ConnectX-7, which gives it 2 x 200 Gbps networking. The 10 gig port is far from the fastest network interface on the Spark.

But can you hook that up to a normal PC?

Yes. Just buy the Mellanox card. We had a bunch of ConnectX 5 hooked up through SFP. Needs cooling but fast.

Re: Nvidia DGX Spark

#96
post #93

Paper launch. The people I know there who I have asked about it haven't even seen one yet

Ordered one in spring. Delivery time was pushed from July to September. Apparently they had a bug in the HDMI output.

That's eerily similar to what happened to Qualcomm's failed Snapdragon X Elite dev kit. That one eventually shipped in small quantities with a Type-C to HDMI dongle in the box to make up for the built-in HDMI port going missing. Then Qualcomm cancelled the whole project and refunded everyone, including people who had already received their hardware.

Re: Nvidia DGX Spark

#97
post #24

The mainstream options seem to be Ryzen AI Max 395+, ~120 tops (fp8?), 128GB RAM, $1999 Nvidia DGX Spark, ~1000 tops fp4, 128GB RAM, $3999 Mac Studio max spec, ~120 tflops (fp16?), 512GB RAM, 3x bandwidth, $9499 DGX Spark appears to potentially offer the most token per second, but less useful/value as everyday pc.

> Ryzen AI Max 395+, ~120 tops (fp8?), 128GB RAM, $1999 Just got my Framework PC last week. It's easy to setup to run LLMs locally - you have to use Fedora 42, though, because it has the latest drivers. It was super easy to get qwen3-coder-30b (8 bit quant) running in LMStudio at 36 tok/sec.

Hi could you share if you get a decent coding performance (quality wise) with this setup? IE. Is it good enough to replace say Claude Code?

Re: Nvidia DGX Spark

#99
post #64

Earlier quoted context omitted.

M4 max has more than double the bandwidth. Strix Halo has the same and I agree it’s overrated.

I would expect/hope that DGX would be able to make better use of its bandwidth than the M4 Max. Will need to wait and see benchmarks.

It should. It has tensor cores which should drastically improve prompt processing. It should also be highly optimized for most AI apps.

Re: Nvidia DGX Spark

#100

Earlier quoted context omitted.

Again, prompt processing isn't the major problem here. It's bandwidth. 256GB/s bandwidth (maybe ~210 in real world) limits the tokens per second well before prompt processing. Not entirely sure how your ARM statement matters here. This is unified memory.

[flagged]

What model are you running?

I suspect that you’re running a very large model like DeepSeek in coherent memory?

Keep in mind that this little DGX only has 128GB which means it can run fairly small models such as qwen3 coder where prompt processing is not an issue.

I’m not doubting your experience with GH200 but it doesn’t seem relevant here because the bandwidth for Spark is the bottleneck well before the prompt processing.

Post reply on HN