Earlier quoted context omitted.
Garbage operating system support. If you can’t do Linux support it’s a bit pointless because there’s two platforms for this that matter: Linux and Darwin. Qualcomm is like AMD was for GPUs for like decades. Lots of announcements and people on the Internet are huge fans based on web pages they’ve read but if you try to make it work it’s a nightmare. Snapdragon X Elite doesn’t work on Linux so it’s a pointless platform…
How does Darwin matter more than Windows?
Nvidia is proposing a beast of a CPU system for Windows PCs
461–470 of 581 posts
Re: Nvidia is proposing a beast of a CPU system for Windows PCs
#462Earlier quoted context omitted.
And conveniently, by making your machine non upgradeable, it allows the manufacturer to enforce market segmentation / charge a huge premium for small RAM upgrade ( a la Apple)
There is LPCAMM2, if manufacturers want to use it. So, it does not have to be soldered.
I don't know how linear or sensitive CPU and GPU benchmarks are to such a 20% slowdown, but i don't think Apple wants to pay it. And it looks like the next generation will be even closer to the SOC.
Re: Nvidia is proposing a beast of a CPU system for Windows PCs
#463Earlier quoted context omitted.
And here I am with 128GB Strix Halo longingly eyeing the Blackwell cards that spit tokens 10-20x the speed. The question is ultimate shape of knowledge compression and bandwidth optimization at which we arrive I suppose.
If you haven't already, check/increase the GPU memory carve-out on your UEFI. More details: https://rocm.docs.amd.com/en/docs-7.2.0/how-to/system-optimi...
Re: Nvidia is proposing a beast of a CPU system for Windows PCs
#464Earlier quoted context omitted.
If this thing only has as much gpu bandwidth as the spark, it’s kinda pointles
Not true. This is aimed squarely at the Strix Halo and Mac markets. It's basically just strictly better than the Strix, and it's not clear cut vs that Macs in any sort of blanket statement. My M5 Max 128gb MBP decodes faster than one of my Sparks, but the Spark's prefill is so much faster it can often answer the same query before the mac's prefill is finished. If you have large prompts, low cacheability, etc., a spar…
Seems niche to be both uncacheable and long context?
Re: Nvidia is proposing a beast of a CPU system for Windows PCs
#465"I am not sure how many people will run AI models locally. It still seems like a niche application to me. However, it will make decent machines to play video games." I don't know who will be the winner but with some of the recent releases from gemma it seems more probable that you may run some models locally if only from a cost perspective, not even considering business security. Not sure how this type of architectur…
> you may run some models locally if only from a cost perspective I have a hard time believing running a model on a laptop will be cheaper than running it in a datacenter. Why wouldn't economies of scale apply here as with every other computation?
Re: Nvidia is proposing a beast of a CPU system for Windows PCs
#466The Unified Memory pool is what will continue to be the “game changer” in systems architecture, especially outside of data centers. The reality is even cutting edge games and consumer workloads don’t actually take full use of the PCIe bandwidth of the GPU or the bandwidth of its GDDR memory. Even local AI use cases don’t substantially or meaningfully benefit from faster memory, at least to average consumers. A unifie…
Re: Nvidia is proposing a beast of a CPU system for Windows PCs
#467This feels fluff to me on the part of the author (whose work I don’t want to trivialize) but I don’t think they’ve actually looked deeper than a paper spec sheet on this. 1. Yes it has the same number of cores as a 5070 mobile. It’s also running at a shared peak of 2/3 the bandwidth and a shared peak of 2/3 the TDP. The GPU by itself will likely perform at half the dedicated units performance 2. Apple may not have SV…
It is absolutely fluff, and the only reason this worthless tweet is on the front page of HN is that this audience has a habit of canonizing certain people, and treating each of their bowel movements as prophetic. Guy suddenly became aware of a chip that the rest of the industry long knew about, seems completely unaware of the competitors, and posts about how it's a BEAST and will be a GAME CHANGER. Like the DGX Spark…
Re: Nvidia is proposing a beast of a CPU system for Windows PCs
#468The Unified Memory pool is what will continue to be the “game changer” in systems architecture, especially outside of data centers. The reality is even cutting edge games and consumer workloads don’t actually take full use of the PCIe bandwidth of the GPU or the bandwidth of its GDDR memory. Even local AI use cases don’t substantially or meaningfully benefit from faster memory, at least to average consumers. A unifie…
And conveniently, by making your machine non upgradeable, it allows the manufacturer to enforce market segmentation / charge a huge premium for small RAM upgrade ( a la Apple)
Re: Nvidia is proposing a beast of a CPU system for Windows PCs
#469Earlier quoted context omitted.
I think much of the difficulty is just that, for example, the 1.8 TB/s of an RTX 5090 is a lot of bandwidth for a game to use. That's over 50,000 4k textures per second at 32bpp.
That sounds like a lot, but: modern renderers do between 20 to 40 passes, many of them in screen space. And each screen space pass typically reads from at least two input images, sometimes 3 or 4 even with optimally packed inputs. At 60fps that can quickly get up to way over 2000 full screen buffer reads per second and more for less than optimal access patterns in some algorithms. That also doesn't account for textur…
Plus, DLSS can greatly reduce the bandwidth requirements for 4K gaming.
Re: Nvidia is proposing a beast of a CPU system for Windows PCs
#470> The game changer is the unified 128 GB memory. That is the path Apple took years ago. Instead of separate memory for the CPU and GPU, everything shares a single pool. It is increasingly popular. > The memory is not as fast as dedicated GPU memory, but it is cheap enough while delivering enough bandwidth to run AI models locally. So, the reason "dedicated GPU memory" is fast, isn't because it's "dedicated"; it's bec…