Nvidia DGX GH200 Whitepaper
resources.nvidia.com
Nvidia DGX GH200 Whitepaper
1–10 of 45 posts
Re: Nvidia DGX GH200 Whitepaper
#2Details on NVLink would be great. For example, the needs and problems solved by their custom cables seemingly required by NVLink would be worth a whitepaper.
Don't get me wrong, this is still great the general public can get a glimpse into Grace Hopper. And they do a good job of simplifying while throwing around mind-boggling numbers (the NVLink bandwidth is insane, though no words on latency, crucial for remote memory access).
Re: Nvidia DGX GH200 Whitepaper
#3Re: Nvidia DGX GH200 Whitepaper
#4Re: Nvidia DGX GH200 Whitepaper
#5So basically 2x faster than H100
Re: Nvidia DGX GH200 Whitepaper
#6Re: Nvidia DGX GH200 Whitepaper
#7So basically 2x faster than H100
> DGX GH200 enables more efficient parallel mapping and alleviates the networking communication bottleneck. As a result, up to 1.5x faster training time can be achieved over a DGX H100-based solution for LLM training at scale.
Re: Nvidia DGX GH200 Whitepaper
#8As context: 1x dgx gh200 has 256x gh200s which each have 1x h100 and 1x grace cpu
Re: Nvidia DGX GH200 Whitepaper
#9https://www.anandtech.com/show/18877/nvidia-grace-hopper-has...
Re: Nvidia DGX GH200 Whitepaper
#10On the LLM frontier, we’re starting to hit the limits of reasoning abilities in the current gen.