Live data from Hacker News

Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

phoronix.com

51–60 of 131 posts

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#51

Earlier quoted context omitted.

How does LPDDR5 (This Xe3P) compare with GDDR7 (Nvidia's flagships) when it comes to inference performance? Local inference is an interesting proposition because today in real life, the NV H300 and AMD MI-300 clusters are operated by OpenAI and Anthropic in batching mode, which slows users down as they're forced to wait for enough similar sized queries to arrive. For local inference, no waiting is required - so you c…

Lpddr5x (not lpddr5) is 10.7 Gbps. Gddr7 is 32 Gbps. So it's going to be slower

Yes but in matrix multiplication there are O(N²) numbers and O(N³) multiplications, so it might be possible that you are bounded by compute speed.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#52

Earlier quoted context omitted.

This is a shareholder “me too” product

What are they gonna do with their own FAB? Not release anything? There'll be a good market share for comparatively "lower power/ good enough" local AI. Check out Alez Ziskind's analysis of the B50 Pro [0]. Intel has an entire line-up of cheap GPUs that perform admirably for local use cases. This guy is building a rack on B580s and the driver update alone has pushed his rig from 30 t/s to 90 t/s. [1] 0: https://www.yo…

Watson…

Yeah even RTX’s are limited in this space due to lack of tensor cores. It’s a race to integrate more cores and faster memory buses. My suspicion is this is more me too product announcement so they can play partner to their business opportunities and continue greasing their wheels.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#54
Between 18A becoming viable and this, it seems Intel is finally climbing out of the hole it's been in for years.

Makes me wonder whether Gelsinger put all this in motion, or if the new CEO lit a fire under everyone. Kinda a shame if it's the former...

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#55

Any business people here that can explain why companies announce products a year before their release? I can understand getting consumers excited but it also tells competitors what you are doing giving them time to make changes of their own. What's the advantage here?

It can also prevent competitors from entering a particular space. I was told as an undergraduate that UNIX was irrelevant because the upcoming Windows NT would be POSIX compliant. It took a _very_ long time before that happened (and for a very flexible version of "compliant"), but the pointy-headed bosses thought that buying Microsoft was the future. And at first glance the upcoming NT _looked_ as if the TCO would be much lower than AIX, HPuX or Solaris.

Then of course Linux took over everywhere except the desktop.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#56
post #36

Earlier quoted context omitted.

DGX Spark is $4000... this might ( might ) not be? (and with more memory)

This starts shipping in 2027. I'm sure you can buy a DGX Spark for less than $4k in 2 years time.

But good luck with Nvidia not turning it into abandoware.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#57
post #23
post #16

I have no idea of the likely price, but (IMO) this is the sort of disruption that Intel needs to aim at if it's going to make some sort of dent in this market. If they could release this for around the price of a 5090, it would be very interesting.

With this much ram don’t expect anything remotely affordable by civilians.

Uncle Sam owns a good chunk of Intel now. "Not affordable by civilians" might be precisely the target market: the DoD/national intelligence agencies have money to burn, can fund things long enough to stabilize Intel a little, and in exchange they get first dibs on everything.

Intel for intel on your Intels, perhaps.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#58
post #55

Any business people here that can explain why companies announce products a year before their release? I can understand getting consumers excited but it also tells competitors what you are doing giving them time to make changes of their own. What's the advantage here?

It can also prevent competitors from entering a particular space. I was told as an undergraduate that UNIX was irrelevant because the upcoming Windows NT would be POSIX compliant. It took a _very_ long time before that happened (and for a very flexible version of "compliant"), but the pointy-headed bosses thought that buying Microsoft was the future. And at first glance the upcoming NT _looked_ as if the TCO would be…

That wasn't even necessarily false. Windows NT on commodity hardware from the likes of Dell arguably did have a lower TCO than proprietary UNIX on proprietary hardware.

But then Linux on that same commodity hardware was lower yet.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#59

Earlier quoted context omitted.

It was many generations before vector operations were moved onto graphics chips.

If you s/graphics/3d graphics does that still hold true?

Yes. The earliest consumer PC 3D graphics cards just rasterized pre-transformed triangles and that's it; the CPU had to do pretty much all the math (but drawing the pixels was considered the hard part). Later, "Hardware Transform and Lighting (T&L)" was introduced circa 2000 by cards like the GeForce 256.

Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM

#60
post #12

Earlier quoted context omitted.

AI is not going anywhere. Now everyone wants to get a piece. Local inference is expected to grow. Documents, image, video, etc processing. Another obvious is driverless farm vehicles and other automated equipment. "Assisted" books, images, news,.. already and grows fast. Translation also a fact.

The technology, maybe - and if on local. The public co valuations of quickly depreciating chip hoarders selling expensive fever dreams to enterprises are gonna pop though. Spend 3-7 USD for 20 cents in return and 95% project failures rates for quarters on end aren't gonna go unnoticed on Wall St.

So far there is no 'plateau' in the nearest future. 'AI' as a science and its applications should develop further for the next several years. Models will get more efficient, but still the bigger the better. This is obvious. Even if models don't scale up well, they can be used collectively in parallel 'brainstorming'. This will still create demand for hardware. Stagnation is still possible in case of recession. In this case even stable businesses will suffer.

As for efficiency, replacing one programmer in group of 10 with AI already will increase productivity and lower the price. In most cases. In reality adding AI accounts to existing group works better. This is _now_, not hopes or sci-fi.

That's why I'm saying there is no way back. 'AI winter' is as likely as smartphones winter.

Post reply on HN