Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
111–120 of 131 posts
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#112Honestly, Intel just has to build a GPU with insane amount of VRAM. It doesn't even have to be the fastest to compete... just a ton of vram for dirt cheap
It’s LPDDR5x It’s gonna be slowwww It’s gonna be what, 273GB/sec vram bandwidth at most? Might as well as buy an AND 395+ 128GB right now for the same inference performance and slightly less VRAM.
If its fast LPDDR5x (9600 MT/s) with 512 bit bus width (8 64bit channels (actually multiples of quad 16 bit subchannel nonsense)) it could be upwards of 600 GB/s. Lots of bandwidth like the beefy macs have.
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#113Earlier quoted context omitted.
And even then, you couldn't really get any sort of serious matmul out of it; they were per-vertex, not per-pixel. Per-pixel matmul (which is what you really need for anything resembling GPGPU) came with Shader Model 2.0, circa 2002; Radeon 9700, the GeForce FX series and the likes. CUDA didn't exist (nor really any other form of compute shaders), but you could wrangle it with pixel shaders, and some of us did.
Oh man, I forgot about doing vector math using OpenGL textures as "hardware acceleration". And it would be many more years before it was reasonable to require a GPU with programmable shaders; having to support fixed-function was a fact of life for most of the 2000's.
Granted, if you didn't have the “squared blend” extension, it would be an approximation, but still a pretty convincing one.
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#114Earlier quoted context omitted.
With 160GB, surely they can add more channels to compensate?
Is there anything preventing them from using heterogeneous memory chips, like 1/4 GDDR7 and 3/4 LPDDR? It could enable new MEO-like architectures with finer-grained performance tuning for long contexts.
If the internal bus architecture is anything similar to QPI, getting the 'different' parts to communicate reliably is probably also a pain.
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#115Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#116I have no idea of the likely price, but (IMO) this is the sort of disruption that Intel needs to aim at if it's going to make some sort of dent in this market. If they could release this for around the price of a 5090, it would be very interesting.
> If they could release this for around the price of a 5090 This is not targeted at consumers. It’s competing with nVidia’s high RAM workstation cards. Think $10K price range, not $1-2K. The 160GB of LPDDR5X chips alone is expensive enough that they couldn’t release this at the $2K price point unless they felt like giving it away (which they don’t)
Any higher and its not really a disruption
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#117Honestly, Intel just has to build a GPU with insane amount of VRAM. It doesn't even have to be the fastest to compete... just a ton of vram for dirt cheap
It’s LPDDR5x It’s gonna be slowwww It’s gonna be what, 273GB/sec vram bandwidth at most? Might as well as buy an AND 395+ 128GB right now for the same inference performance and slightly less VRAM.
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#118Earlier quoted context omitted.
So far there is no 'plateau' in the nearest future. 'AI' as a science and its applications should develop further for the next several years. Models will get more efficient, but still the bigger the better. This is obvious. Even if models don't scale up well, they can be used collectively in parallel 'brainstorming'. This will still create demand for hardware. Stagnation is still possible in case of recession. In thi…
Your entire argument is whoefully ignoring the CapEx economics of all of this. But that's the foundation. And there is a plateau in real money spent on AI chips. You're ignoring a whole group of economic and finance professionals as well as - if you're inclined to listen to their voices more - Sama calling it a bubble. If not for AI spending, the US already would be in a recession. So your argument might sound nice a…
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#119Earlier quoted context omitted.
160 GB LPDDR5 is ~$1,200 retail so the card could be sold for $2,000. The price will depend on how desperate Intel is. Intel probably can't copy Nvidia's pricing.
It’d be a disaster for Intel if it sold for less than 3k, personally I think they’re aiming for break even at 5k a pop at least, and I wouldn’t be surprised to advertise 2x memory at half nvidia price, which would put it at ~15-20k? and a healthy margin which they need like oxygen now. Of course it’s all for naught if it doesn’t perform compute-wise.
Re: Intel Announces Inference-Optimized Xe3P Graphics Card with 160GB VRAM
#120What price is this sitting at? Because if its software support is decent then Intel might have just managed to break into the hardware for AI on the edge. Examples like self hosted LLM finetuning and RAG on a old dell or HP server with these type of cards on them.
> Examples like self hosted LLM finetuning and RAG on an old dell or HP server with these type of cards on them. This won’t be in the price range of an old Dell server or a fun impulse buy for a hobbyist. 160GB of raw LPDDR5X chips alone is not cheap. This is a server/workstation grade card and the price is going where the market will allow. Consider that an nVidia card with almost half the RAM is going to cost $8K o…