Live data from Hacker News

Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

phoronix.com

31–40 of 131 posts

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#32
post #31
post #17

A not-absurdly-priced card that can run big models (even quantized) would sell like crazy. Lots and lots of fast RAM is key.

Isn't that precisely what DGX Spark is designed for? How is this better?

DGX Spark is $4000... this might (might) not be? (and with more memory)

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#33
post #17

A not-absurdly-priced card that can run big models (even quantized) would sell like crazy. Lots and lots of fast RAM is key.

How does LPDDR5 (This Xe3P) compare with GDDR7 (Nvidia's flagships) when it comes to inference performance? Local inference is an interesting proposition because today in real life, the NV H300 and AMD MI-300 clusters are operated by OpenAI and Anthropic in batching mode, which slows users down as they're forced to wait for enough similar sized queries to arrive. For local inference, no waiting is required - so you c…

I asked GPT to pull real stats on both. Looks like the 50-series RAM is about 3X that of the Xe3P, but it wanted to remind me that this new Intel card is designed for data centers and is much lower power, and that the comparable Nvidia server cards (e.g. H200) have even better RAM than GDDR7, so the difference would be even higher for cloud compute.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#34
post #29

Funny they still call them graphics cards when they're really... I dont know, matmul cards ? Tensor cards ? TPU ? Well that sums it up maybe, what those are are really CUDA cards.

Dude, this is asinine. Graphics cards have been doing matrix and vector operations since they were invented. No one had a problem with calling matrix multiplers graphics cards until it became cool to hate AI.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#35
post #17

A not-absurdly-priced card that can run big models (even quantized) would sell like crazy. Lots and lots of fast RAM is key.

How does LPDDR5 (This Xe3P) compare with GDDR7 (Nvidia's flagships) when it comes to inference performance? Local inference is an interesting proposition because today in real life, the NV H300 and AMD MI-300 clusters are operated by OpenAI and Anthropic in batching mode, which slows users down as they're forced to wait for enough similar sized queries to arrive. For local inference, no waiting is required - so you c…

Lpddr5x (not lpddr5) is 10.7 Gbps. Gddr7 is 32 Gbps. So it's going to be slower

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#36
post #31

Earlier quoted context omitted.

Isn't that precisely what DGX Spark is designed for? How is this better?

DGX Spark is $4000... this might ( might ) not be? (and with more memory)

This starts shipping in 2027. I'm sure you can buy a DGX Spark for less than $4k in 2 years time.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#37
post #29

Funny they still call them graphics cards when they're really... I dont know, matmul cards ? Tensor cards ? TPU ? Well that sums it up maybe, what those are are really CUDA cards.

Dude, this is asinine. Graphics cards have been doing matrix and vector operations since they were invented. No one had a problem with calling matrix multiplers graphics cards until it became cool to hate AI.

It was many generations before vector operations were moved onto graphics chips.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#38
post #22

Earlier quoted context omitted.

There is a serious possibility this isn’t a bubble. Too many people watched the big short and now call every bull a bubble; maybe the bubble was the dollar and it’s popping now instead.

Have you looked in detail at the economics of this? Career finance professionals are calling it a bubble, not due to their suddenly found deep technological expertise, but because public cos like FAANG et. al are engaging in typical bubble like behavior: Shifting capex away from their balance sheets into SPACs co-financed by private equity. This is not a consumer debt bubble, it's gonna be a private market bubble. Bu…

It's possible, circular financing is definitely fishy, but OTOH every openai deal sama makes is swallowed by willing buyers at a fair market price. We'll be in a bubble when all the bears are dead and everyone accepts 'a new paradigm', not before; there's plenty of upside capitulation left judging by some hedge fund returns this year.

...and again, this is assuming AI capability stops growing exponentially in the widest possible sense (today, 50%-task-completion time horizon doubles ~7 months).

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#39
post #30

It'll be either "cheap" like the DGX Spark (with crap memory bandwidth) or overpriced with the bus width of a M4 Max with the rhetoric of Intel's 50% margin.

Or it will be cheap, with the ability to expand 8X on a server. Particularly with PCIe 6.0 coming soon, might be a very attractive package.

https://www.linkedin.com/posts/storagereview_storagereview-a...

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#40

Earlier quoted context omitted.

Dude, this is asinine. Graphics cards have been doing matrix and vector operations since they were invented. No one had a problem with calling matrix multiplers graphics cards until it became cool to hate AI.

It was many generations before vector operations were moved onto graphics chips.

If you s/graphics/3d graphics does that still hold true?
Post reply on HN