Earlier quoted context omitted.
Because Apple isn't playing the same game as everyone else. They have the money and clout to buy out TSMCs bleeding-edge processes and leave everyone else with the scraps, and their silicon is only sold in machines with extremely fat margins that can easily absorb the BOM cost of making huge chips on the most expensive processes money can buy.
Bleeding edge processes is what Intel specializes in. Unlike Apple, they don’t need TSMC. This should have been a huge advantage for Intel. Maybe that’s why Gelsinger got the boot.
Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
281–290 of 722 posts
Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
#282Earlier quoted context omitted.
The M4 Max needs an enormous 512bit memory bus to extract enough bandwidth out of those LPDDR5x chips, while the GPUs that Intel just launched are 192/160bit and even flagships rarely exceed 384bit. They can't just slap more memory on the board, they would need to dedicate significantly more silicon area to memory IO and drive up the cost of the part, assuming their architecture would even scale that wide without hit…
Man, I'm old enough to remember when 512 was a thing for consumer cards back when we had 4-8gb memory Sure that was only gddr5 and not gddr6 or lpddr5, but I would have bet we'd be up to 512bit again 10 years down the line.. (I mean supposedly hbm3 has done 1024-2048bit busses but that seems more research or super high end cards, not consumer)
Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
#283Earlier quoted context omitted.
> ML is about hit another winter. I agree ML is about to hit (or has likely already hit) some serious constraints compared to breathless predictions of two years ago. I don't think there's anything equivalent to the AI winter on the horizon, though—LLMs even operated by people who have no clue how the underlying mechanism functions are still far more empowered than anything like the primitives of the 80s enabled.
What we had in the 80s was barely able to perform spell-check, free downloadable LLMs today are mind-blowing even in comparison to GPT-2.
Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
#284Earlier quoted context omitted.
This is the most script kiddy comment I've seen in a while. llama.cpp is just inference, not training, and the CUDA backend is still the fastest one by far. No one is even close to matching CUDA on either training or inference. The closest is AMD with ROCm, but there's likely a decade of work to be done to be competitive.
Yes, and inference is a huge market in itself and potentially larger than training (gut feeling haven’t run numbers) Keep NVIDIA for training and Intel/AMD/Cerebras/… for interference.
Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
#285Earlier quoted context omitted.
The M4 Max needs an enormous 512bit memory bus to extract enough bandwidth out of those LPDDR5x chips, while the GPUs that Intel just launched are 192/160bit and even flagships rarely exceed 384bit. They can't just slap more memory on the board, they would need to dedicate significantly more silicon area to memory IO and drive up the cost of the part, assuming their architecture would even scale that wide without hit…
Apple could do it. Why can’t Intel?
Everyone else wants configurable RAM that scales both down (to 16GB) and up (to 2TB), to cover smaller laptops and bigger servers.
GPUs with soldered on RAM has 500GB/sec bandwidths, far in excess of Apples chips. So the 8GB or 16GB offered by NVidia or AMD is just far superior at vid o game graphics (where textures are the priority)
Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
#286Why don't they just release a basic GPU with 128GB RAM and eat NVidia's local generative AI lunch? The networking effect of all devs porting their LLMs etc. to that card would instantly put them as a major CUDA threat. But beancounters running the company would never get such an idea...
Disclosure: HPC admin who works with NIVIDA cards here. Because, no. It's not as simple as that. NVIDIA has a complete ecosystem now. They have cards. They have cards of cards (platforms), which they produce, validate and sell. They have NVLink crossbars and switches which connects these cards on their card of cards with very high speeds and low latency. For inter-server communication they have libraries which coordi…
Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
#287Earlier quoted context omitted.
You're asking for a GPU die at least as large as NVIDIA's TU102 that was $1k in 2018 when paired with only 11GB of RAM (because $1k couldn't get you a fully-enabled die to use 12GB of RAM). I think you're off by at least a factor of two in your cost estimates.
Intel has Xeon Phi which was a spin-off of their first attempt at GPU so they have a lot of tech in place they can reuse already. They don't need to go with GDDRx/HBMx designs that require large dies.
Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
#288Why don't they just release a basic GPU with 128GB RAM and eat NVidia's local generative AI lunch? The networking effect of all devs porting their LLMs etc. to that card would instantly put them as a major CUDA threat. But beancounters running the company would never get such an idea...
Disclosure: HPC admin who works with NIVIDA cards here. Because, no. It's not as simple as that. NVIDIA has a complete ecosystem now. They have cards. They have cards of cards (platforms), which they produce, validate and sell. They have NVLink crossbars and switches which connects these cards on their card of cards with very high speeds and low latency. For inter-server communication they have libraries which coordi…
Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
#289Earlier quoted context omitted.
Apple could do it. Why can’t Intel?
Because LPDDR5x is soldered on RAM. Everyone else wants configurable RAM that scales both down (to 16GB) and up (to 2TB), to cover smaller laptops and bigger servers. GPUs with soldered on RAM has 500GB/sec bandwidths, far in excess of Apples chips. So the 8GB or 16GB offered by NVidia or AMD is just far superior at vid o game graphics (where textures are the priority)
Apple is doing 800GB/sec on the M2 Ultra and should reach about 1TB/sec with the M4 Ultra, but that's still lagging behind GPUs. The 4090 was already at the 1TB/sec mark two years ago, the 5090 is supposedly aiming for 1.5TB/sec, and the H200 is doing 5TB/sec.
Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support
#290Earlier quoted context omitted.
What is the newest platform that lacks resizable BAR? It was standardized in 2006. Is 4060-level graphics performance useful in whatever old computer has that problem?
Sandy Bridge (2009) is still a very usable CPU with a modern GPU. In theory Sandy Bridge supported resizable BAR but in practice they didn't. I think the problem was BIOS's.