Earlier quoted context omitted.
How does LPDDR5 (This Xe3P) compare with GDDR7 (Nvidia's flagships) when it comes to inference performance? Local inference is an interesting proposition because today in real life, the NV H300 and AMD MI-300 clusters are operated by OpenAI and Anthropic in batching mode, which slows users down as they're forced to wait for enough similar sized queries to arrive. For local inference, no waiting is required - so you c…
Lpddr5x (not lpddr5) is 10.7 Gbps. Gddr7 is 32 Gbps. So it's going to be slower
Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
51–60 of 131 posts
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#52Earlier quoted context omitted.
This is a shareholder “me too” product
What are they gonna do with their own FAB? Not release anything? There'll be a good market share for comparatively "lower power/ good enough" local AI. Check out Alez Ziskind's analysis of the B50 Pro [0]. Intel has an entire line-up of cheap GPUs that perform admirably for local use cases. This guy is building a rack on B580s and the driver update alone has pushed his rig from 30 t/s to 90 t/s. [1] 0: https://www.yo…
Yeah even RTX’s are limited in this space due to lack of tensor cores. It’s a race to integrate more cores and faster memory buses. My suspicion is this is more me too product announcement so they can play partner to their business opportunities and continue greasing their wheels.
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#53Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#54Makes me wonder whether Gelsinger put all this in motion, or if the new CEO lit a fire under everyone. Kinda a shame if it's the former...
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#55Any business people here that can explain why companies announce products a year before their release? I can understand getting consumers excited but it also tells competitors what you are doing giving them time to make changes of their own. What's the advantage here?
Then of course Linux took over everywhere except the desktop.
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#56Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#57I have no idea of the likely price, but (IMO) this is the sort of disruption that Intel needs to aim at if it's going to make some sort of dent in this market. If they could release this for around the price of a 5090, it would be very interesting.
With this much ram don’t expect anything remotely affordable by civilians.
Intel for intel on your Intels, perhaps.
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#58Any business people here that can explain why companies announce products a year before their release? I can understand getting consumers excited but it also tells competitors what you are doing giving them time to make changes of their own. What's the advantage here?
It can also prevent competitors from entering a particular space. I was told as an undergraduate that UNIX was irrelevant because the upcoming Windows NT would be POSIX compliant. It took a _very_ long time before that happened (and for a very flexible version of "compliant"), but the pointy-headed bosses thought that buying Microsoft was the future. And at first glance the upcoming NT _looked_ as if the TCO would be…
But then Linux on that same commodity hardware was lower yet.
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#59Earlier quoted context omitted.
It was many generations before vector operations were moved onto graphics chips.
If you s/graphics/3d graphics does that still hold true?
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#60Earlier quoted context omitted.
AI is not going anywhere. Now everyone wants to get a piece. Local inference is expected to grow. Documents, image, video, etc processing. Another obvious is driverless farm vehicles and other automated equipment. "Assisted" books, images, news,.. already and grows fast. Translation also a fact.
The technology, maybe - and if on local. The public co valuations of quickly depreciating chip hoarders selling expensive fever dreams to enterprises are gonna pop though. Spend 3-7 USD for 20 cents in return and 95% project failures rates for quarters on end aren't gonna go unnoticed on Wall St.
As for efficiency, replacing one programmer in group of 10 with AI already will increase productivity and lower the price. In most cases. In reality adding AI accounts to existing group works better. This is _now_, not hopes or sci-fi.
That's why I'm saying there is no way back. 'AI winter' is as likely as smartphones winter.