Live data from Hacker News

Nvidia Hopper GPU Architecture and H100 Accelerator

anandtech.com

151–160 of 183 posts

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#152

Earlier quoted context omitted.

All the AI software running on these data-centre chips is almost exclusively running on Linux. I wish people would stop talking rubbish about NVIDIA's Linux support.

That's because nvidias linux support for consumers is indeed trash, while their creators/business/creatives software (eg CUDA) is not trash, but you mostly hear consumers trashing nvidia.

Only FOSS zealots actually, the rest of us is quite ok with their binary drivers.

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#154

How does a DGX Pod w/ the new 3.2Tbps per machine NVLINK switch compare to Tesla Dojo?

Tesla Dojo Training tile (25x D1): 565 TF FP32 / 9 PF BF16/CFP8 / 11GB SRAM / 10kW

NVIDIA DGX H100 (8x H100): 480 TF FP32 / 8 PF+ TF16 / 16 PF INT8 / 640GB HBM3 / 10kW

Dojo off-chip BW: 16 TB/s / 36TB/s off-tile

H100 off-chip BW: 3.9TB/s / 400GB/s off-DGX

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#156
post #126

Earlier quoted context omitted.

>you don't need to remember Intel codenames because a Core 12700 is obviously newer than a Core 11700 J3710, 7th Gen J3060, 8th Gen J4205, 8th Gen J4125, 9th Gen i3-5005U, 5th Gen N5095, 10th Gen i7-3770, 3rd Gen 3865U, 7th Gen N3060, 8th Gen

And an AMD 5700U is older than a 5400U as well. A 3400G is older than a 3100X. 3300X isn't really distinctive from 3100X, both are quad-core configurations (but different CCD/cache configurations, which is of course the name doesn't really disclose to the consumer). It happens, naming is a complex topic and there's a lot of dimensions to a product. In general, complaining about naming is peak bikeshedding for the tec…

yes, it's a bit of a shitshow, as mutually evidenced. unless consumers brush up on such intricate details (most do not), they will inevitably fall into traps such as "i7 is better than i3" e.g. i7-2600 being outperformed by i3-10100 and "quad core is better than dual core". marketing is becoming more focused on generations now which is a prudent move: "10th Gen is better than 2nd Gen" but it will be at least a decade before the shitshow is swept

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#157

Earlier quoted context omitted.

Yes, nor Lambda Labs or Exxact Corporation have them available last time I checked (last week). Both citing high demand as the reason for it being unavailable.

We (Lambda) have all of the different NVIDIA GPUs in stock ---- can you send a message to sales@lambdalabs.com and check in again with your requirements? We're seeing a lot more stock these days as the supply chain crisis of 2021 comes to an end.

I talked with you (Lambda Labs) just a week ago about the A100 specifically and you said that the demand was higher than the supply, and that people should check once a day or something like that to see if it's available in your dashboard. If you clearly have it available now, please say so outright instead of trying to push some other offer on me in emails :)

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#158
post #100

Earlier quoted context omitted.

Most of the propane heater is a fan in a tube, the flame is probably quite smaller than a CPU package.

I've got an 8kW wood stove and that thing gets rather hot to touch - as in, you will get a blister... 40kW is a small city car worth of power.

If you think about it, cars can manage cooling 200+kW with a radiator - and a lot of airflow.

That's still an amazing amount of power though. I can't help thinking about the kind of sear you could get on a steak with 40kW of power :D

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#159
post #151

Anyone find details about the DPX instructions for dynamic programming?

There are some deep dive sessions at GTC that will probably go into them.

The NVIDIA statement about it:

> For now, DPX ISA details are available to early access partners. We anticipate broader info availability aligned with CUDA 12.0 release later this year.

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#160

The Tensor cores will be great for machine learning and the FP32/FP64 fantastic for HPC, but I'd be surprised if there were a lot of applications using both of these features at once. I wonder if there's room for a competitor to come in and sell another huge accelerator but with only one of these two features either at a lower price or with more performance? Perhaps the power density would be too high if everything w…

Just look at Cerebras Wafer-Scale Engine it's basically a chip the size of a Wafer but it's not cheap...
Post reply on HN