Live data from Hacker News

Nvidia Hopper GPU Architecture and H100 Accelerator

anandtech.com

131–140 of 183 posts

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#132

"Combined with the additional memory on H100 and the faster NVLink 4 I/O, and NVIDIA claims that a large cluster of GPUs can train a transformer up to 9x faster, which would bring down training times on today’s largest models down to a more reasonable period of time, and make even larger models more practical to tackle." Looking good.

[deleted]

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#133

This would be cool if they had decent drivers for Linux.

They do have good drivers for Linux for the things this chips is intended to be used for (research, ML).

If you haven't had any issues with NVIDIA Linux drivers, you can count yourself extremely lucky. In the past, I had a 50/50 chance of boot failure after installing CUDA drivers over 12 different systems. Mainline Ubuntu drivers are somewhat stable, but installing a specific CUDA version from the official NVIDIA repos rarely works on the first try. Switching from Tensorflow to PyTorch has helped a lot though, as Tensorflow was much more picky about the installed CUDA version.

Obligatory Linus Torvals on NVIDIA: https://www.youtube.com/watch?v=_36yNWw_07g

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#134

Off topic but I can't stand when corporations use actual people's names for their marketing who never gave them the permission to do so. For something like Shakespeare or Cicero I'm OK with it but Grace Hopper was alive in my lifetime, and even Tesla feels a little weird. What gives you the right to use that person's reputation to shill your product?

I generally agree with you, but in this case I suspect Grace Hopper would be honored by it and also impressed with the engineering here. It's not like they slapped her name on a soda can or something.

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#135
post #81

Earlier quoted context omitted.

Did you check Lambda or Exxact?

Yes, nor Lambda Labs or Exxact Corporation have them available last time I checked (last week). Both citing high demand as the reason for it being unavailable.

We (Lambda) have all of the different NVIDIA GPUs in stock ---- can you send a message to sales@lambdalabs.com and check in again with your requirements? We're seeing a lot more stock these days as the supply chain crisis of 2021 comes to an end.

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#137
post #34
post #3

Sounds like we need some new training methods. If training could take place locally and asynchronously instead of globally through backpropagation, the amount of energy could probably be significantly reduced.

Trying to reduce energy consumption for ML like this is so silly.

Reducing energy consumption for computation is not silly.

We're at a point we we're turning into a computation driven society and computation is becoming a globally relevant power consumption aspect.

> global data centers likely consumed around 205 terawatt-hours (TWh) in 2018, or 1 percent of global electricity use

And that's just data centers, if you add all client devices you probably double that.

Plus that number will only continue to grow.

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#138

Earlier quoted context omitted.

All the AI software running on these data-centre chips is almost exclusively running on Linux. I wish people would stop talking rubbish about NVIDIA's Linux support.

That's because nvidias linux support for consumers is indeed trash, while their creators/business/creatives software (eg CUDA) is not trash, but you mostly hear consumers trashing nvidia.

They don't make (relevant) money from consumer hardware on Linux.

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#139
post #96

Earlier quoted context omitted.

Your kids have the right to everything you own ( including your name) by default unless you take steps to change that, say using a will or estate.

Yes, I know, I'm saying that it should not be that way. Rights to your likeness should end at your death unless you specifically write down otherwise.

Do you have kids?

Re: Nvidia Hopper GPU Architecture and H100 Accelerator

#140

So the product naming for Nvidia's server-GPUs by compute power now goes: P100 -> V100 -> A100 -> H100 This is not confusing at all.

I think this is less of an issue since these GPUs are not meant for the everyman, so basically the handful of server integrators can figure this out by themselves.

And for your typical dev - they'll interact with the GPU through a cloud provider, where they can easily know that a G5 instance is newer than a G4 one.

Post reply on HN