Live data from Hacker News

Intel Gaudi 3 AI Accelerator

intel.com

71–80 of 260 posts

Re: Intel Gaudi 3 AI Accelerator

#71

https://www.merriam-webster.com/dictionary/gaudy

Honestly, I thought the same thing upon reading the name. I'm aware of the reference to Antoni Gaudí, but having the name sound so close to gaudy seems a bit unfortunate. Surely they must've had better options? Then again I don't know how these sorts of names get decided anymore.

to be fair intel is not known for naming things well.

Re: Intel Gaudi 3 AI Accelerator

#72
post #53

This is a bit snarky — but will Intel actually keep this product line alive for more than a few years? Having been bitten by building products around some of their non-x86 offerings where they killed good IP off and then failed to support it… I’m skeptical. I truly do hope it is successful so we can have some alternative accelerators.

Itanic was a fun era

Re: Intel Gaudi 3 AI Accelerator

#73
I feel a little misled by the speedup numbers. They are comparing lower batch size h100/200 numbers to higher batch size gaudi 3 numbers for throughput (which is heavily improved by increasing batch size). I feel like there are some inference scenarios where this is better, but its really hard to tell from the numbers in the paper.

Re: Intel Gaudi 3 AI Accelerator

#74

Earlier quoted context omitted.

[flagged]

Effectively everyone that has attempted to use AMD hardware for ML comes away with these opinions, the main difference is how angrily they express it.

We are the 4th non-hyperscaler business on the planet to even get access to MI300x and we just got it in early March. From what I understand, hyperscalers have had fantastic uptake of this hardware.

I find it hard to believe "everyone" comes away with these opinions.

https://www.evp.cloud/post/diving-deeper-insights-from-our-l...

Re: Intel Gaudi 3 AI Accelerator

#75
post #37
post #25

Earlier quoted context omitted.

Does it not work for them? Where can I learn why?

Just go have a look around Github issues in their ROCm repositories on Github. A few months back the top excuse re: AMD was that we're not supposed to use their "consumer" cards, however the datacenter stuff is kosher. Well, guess what, we have purchased their datacenter card, MI50, and it's similarly screwed. Too many bugs in the kernel, kernel crashes, hangs, and the ROCm code is buggy / incomplete. When it works,…

Yeah, this has stopped me from trying anything with them. They need to lead with their consumer cards so that developers can test/build/evaluate/gain trust locally and then their enterprise offerings need to 100% guarantee that the stuff developers worked on will work in the data center. I keep hoping to see this but every time I look it isn't there. There is way more support for apple silicon out there than ROCm and that has no path to enterprise. AMD is missing the boat.

Re: Intel Gaudi 3 AI Accelerator

#76
post #50
post #15

Earlier quoted context omitted.

Which would be impressive had it _actually_ worked for ML workloads.

There's a number of scaled AMD deployments, including Lamini ( https://www.lamini.ai/blog/lamini-amd-paving-the-road-to-gpu... ) specifically for LLM's. There's also a number of HPC configurations, including the world's largest publicly disclosed supercomputer (Frontier) and Europe's largest supercomputer (LUMI) running on MI250x. Multiple teams have trained models on those HPC setups too. Do you have any more eviden…

> Do you have any more evidence as to why these categorically don't work?

They don't. Loud voices parroting George, with nothing to back it up.

Here are another couple good links:

https://www.evp.cloud/post/diving-deeper-insights-from-our-l...

https://www.databricks.com/blog/training-llms-scale-amd-mi25...

Re: Intel Gaudi 3 AI Accelerator

#78

A bit surprised that they're using HBM2e, which is what Nvidia A100 (80GB) used back in 2020. But Intel is using 8 stacks here, so Gaudi 3 achieves comparable total bandwidth (3.7TB/s) to H100 (3.4TB/s) which uses 5 stacks of HBM3. Hopefully the older HBM has better supply - HBM3 is hard to get right now! The Gaudi 3 multi-chip package also looks interesting. I see 2 central compute dies, 8 HBM die stacks, and then 6…

> A bit surprised that they're using HBM2e, which is what Nvidia A100 (80GB) used back in 2020. This is one of the secret recipes of Intel. They can use older tech and push it a little further to catch/surpass current gen tech until current gen becomes easier/cheaper to produce/acquire/integrate. They have done it with their first quad core processors by merging two dual core processors (Q6xxx series), or by creating…

Oh dear, Q6600 was so bad, I regret ever owning it

Re: Intel Gaudi 3 AI Accelerator

#79

Earlier quoted context omitted.

Effectively everyone that has attempted to use AMD hardware for ML comes away with these opinions, the main difference is how angrily they express it.

So who's buying all the MI300s? Groq seems to be fine with AMD.

> So who's buying all the MI300s?

We are. Just closed additional funding and getting quotes now.

Re: Intel Gaudi 3 AI Accelerator

#80
post #67

Earlier quoted context omitted.

> A bit surprised that they're using HBM2e, which is what Nvidia A100 (80GB) used back in 2020. This is one of the secret recipes of Intel. They can use older tech and push it a little further to catch/surpass current gen tech until current gen becomes easier/cheaper to produce/acquire/integrate. They have done it with their first quad core processors by merging two dual core processors (Q6xxx series), or by creating…

Interesting. Would you say this means Intel is "back," or just not completely dead?

No, this means Intel has woken up and trying. There's no guarantee in anything. I'm more of an AMD person, but I want to see fierce competition, not monopoly, even if it's "my team's monopoly".
Post reply on HN