Earlier quoted context omitted.
You just blew my mind. That did not occur to me, but it is obvious in retrospect.
Sure, but the way Nvidia names generations is far from obvious. It seems to be “names of famous scientists, progressing in alphabetical order, we skip some letters if we can’t find a well known scientist with a matching last name and are excited about a scientist 2 letters from now, we wrap around to the beginning of the alphabet when we get to the end, and we just skipped from A to H, so expect another wraparound in…
Nvidia Hopper GPU Architecture and H100 Accelerator
41–50 of 183 posts
Re: Nvidia Hopper GPU Architecture and H100 Accelerator
#42From the nvidia page, > 80 billion transistors > Hopper H100 .. generational leap > 9x at-scale training performance over A100 > 30x LLM inference throughput > Transformer Engine .. speed .. 6x without losing accuracy So another monster chip - same size of the Apple M1-max thingy .. I guess it comes down to pricing. The A100 is already ridiculously expensive at $10K. They can this one at $50K and it would sell out?
Re: Nvidia Hopper GPU Architecture and H100 Accelerator
#431000 TFLOPS so i can run my GPT3 in under 100 ms locally :D If 1000 TFLOPS is possible to do in inference time then im speechless
But keep in mind the model won't fit on a single H100 (80GB) because it's 175B params, and ~90GB even with sparse FP8 model weights, and then more needed for live activation memory. So you'll still want atleast 2+ H100s to run inference, and more realistically you would rent a 8xH100 cloud instance.
But yeah the latency will be insanely fast given how massive these models are!
Re: Nvidia Hopper GPU Architecture and H100 Accelerator
#44Earlier quoted context omitted.
Yeah it is, but unless you've memorised the history of Nvidia architectures it doesn't tell you which is the newer one Fermi -> Kepler -> Maxwell -> Pascal -> Volta (HPC only) -> Turing -> Ampere -> Hopper (HPC only?) -> Lovelace?
Isn't this the norm? Only AMD started the trend of naming the uArch with Numbers as Zen 4 or RDNA 3 fairly recently. With Intel it is Haswell > Broadwell > ..... Whatever Lake.
Nvidia's accelerators are just called 100 every time so if you don't remember the order of the letters it's not obvious
They could have just named them P100, V200, A300 and H400 instead
Re: Nvidia Hopper GPU Architecture and H100 Accelerator
#45Earlier quoted context omitted.
That's a strange statement. The vast vast majority of these cards will be in systems running Linux.
I for one suffer deeply when I try to install the nvidia drivers on Linux. The website binaries _always_ break my system Only the ppas from graphics-drivers work properly My experience on windows is much more automatic and it never breaks anything. But I'd rather pay the price (installing on Linux) to avoid windows at all costs
Re: Nvidia Hopper GPU Architecture and H100 Accelerator
#46Re: Nvidia Hopper GPU Architecture and H100 Accelerator
#47This seems fast... TF32 ....... 1,000 TFLOPS (tensor core) FP64/FP32 ... 60 TFLOPS I am more interested in the 144-core Grace CPU Superchip. nVidia is getting into the CPU business...
Re: Nvidia Hopper GPU Architecture and H100 Accelerator
#48Re: Nvidia Hopper GPU Architecture and H100 Accelerator
#491000 TFLOPS so i can run my GPT3 in under 100 ms locally :D If 1000 TFLOPS is possible to do in inference time then im speechless
At inference time it will be possible to do 4000 TFLOPS using sparse FP8 :) But keep in mind the model won't fit on a single H100 (80GB) because it's 175B params, and ~90GB even with sparse FP8 model weights, and then more needed for live activation memory. So you'll still want atleast 2+ H100s to run inference, and more realistically you would rent a 8xH100 cloud instance. But yeah the latency will be insanely fast…
Sounds doable in a generation or two.
Re: Nvidia Hopper GPU Architecture and H100 Accelerator
#50Earlier quoted context omitted.
Don't they provide them for their consumer cards too, just that it's a closed source binary blob?
And not just Linux: FreeBSD. * https://www.nvidia.com/en-us/drivers/unix/freebsd-x64-archiv... * https://www.freshports.org/x11/nvidia-driver Heck, Solaris : * https://www.nvidia.com/en-us/drivers/unix/solaris-display-ar... * https://www.nvidia.com/en-us/drivers/unix/