Live data from Hacker News

Nvidia RTX Spark

nvidia.com

141–150 of 437 posts

Re: Nvidia RTX Spark

#141

This may finally be the chip family ARM on Windows has always needed. Qualcomm's chips have always been dogs with slow off-the-shelf ARM CPU cores that have pathetic single-threaded performance compared to x86 AMD/Intel or ARM Apple Silicon designs.

Qualcomm Snapdragon x1 and upcoming x2 use their Oryon core and have much faster single-thread performance than Intel/Amd and this nvidia soc that uses off-the-shelf arm cores

Re: Nvidia RTX Spark

#142
post #98
post #92

I’m getting more and more convinced that we will end up running LLMs in our personal computers. Which makes me wonder where Anthropic/OpenAIs moats will come from.

Convince me 1. in order to run LLMs, especially the best ones, you need complicated devices which are expensive 2. if you buy one for your personal use, you are probably not going to utilize it all the time and it will be idle a lot It seems to me that it will always be more economical that the LLM-running devices are in a datacenter where it is easier to make sure they are always utilized

3. If your device run on battery, why not using a relatively cheap network call in place of a very power hungry local inference call?

Re: Nvidia RTX Spark

#143
post #27

Unified RAM means its soldered to the mainboard, right? I'm not sure if I like this. Sure for a laptop this might be not a big problem but if this ARM ecosystem is a success it will spread to desktop computers and I fear we could lose the existing modularity.

No, but LPDDR means soldered, there are no LPDDR dimms

There's LPCAMM2, but it's very recent. The Framework Pro laptop supports it, for example, although only on the Intel variant.

Re: Nvidia RTX Spark

#144
post #34
post #7

Earlier quoted context omitted.

yes, same chip + Windows + Screen - ConnectX-7 Smart NIC

What about the desktop version? It seemed like it is not a dgx since it has the CPUs cores done by mediatek

The DGX Spark/GB10 has CPU cores from Mediatek (in a pretty odd cluster configuration, too).

Re: Nvidia RTX Spark

#145

Earlier quoted context omitted.

No. You can get a PowerBook today with 128 GB ram. https://www.bhphotovideo.com/c/product/1957120-REG/apple_mbp...

Or get an AMD 395 laptop or mini PC for half the price of an equivalent mac device

https://onexplayerstore.com/products/onexplayer-super-x?vari...

$3649 with 128GB of ram

Re: Nvidia RTX Spark

#146
post #92

I’m getting more and more convinced that we will end up running LLMs in our personal computers. Which makes me wonder where Anthropic/OpenAIs moats will come from.

We're hitting the atomic limits of what's possible with minimum feature size in silicon. It's also very hard to remove 1 kW of heat from a laptop, let alone do it quietly or on battery.

Re: Nvidia RTX Spark

#147
post #139

What is this product anyway? Is it a general purpose CPU or is it specifically designed for MS Windows? Nvidia stepping back from the open source? "Introducing the NVIDIA RTX Spark™ Superchip. The fusion of NVIDIA AI and RTX graphics in a single chip redefines Windows PCs and delivers amazing creating, AI development, and gaming—on the slimmest, most beautiful RTX laptops ever and small, ultra-efficient desktops."

It’s nivdia attempting to compete with Apple’s M-series

Re: Nvidia RTX Spark

#148

Earlier quoted context omitted.

No. You can get a PowerBook today with 128 GB ram. https://www.bhphotovideo.com/c/product/1957120-REG/apple_mbp...

Or get an AMD 395 laptop or mini PC for half the price of an equivalent mac device

https://www.bosgamepc.com/products/bosgame-m5-ai-mini-deskto...

Bosgame M5 AI Mini Desktop Ryzen AI Max+ 395 96GB variant €1.800,95 (sold out)

128GB+2TB variant €2.401,95 (in stock)

I have the latter, it's fantastic

Re: Nvidia RTX Spark

#150
post #133

Earlier quoted context omitted.

If the workload is offloaded to the chip, why would the host platform matter?

Lots of machine learning workflows support Linux better than Windows, if they run on Windows at all. (e.g. https://docs.vllm.ai/en/latest/getting_started/quickstart/ ) DGX Spark runs Linux, and nobody is going to install Windows on that machine. This laptop got it backwards. If someone decides to run Ollama for local inference with this laptop, they fit perfectly into the "has too much money to waste" bracket, which…

vllm-windows works well enough
Post reply on HN