Live data from Hacker News

Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

phoronix.com

61–70 of 722 posts

Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

#61

Why don't they just release a basic GPU with 128GB RAM and eat NVidia's local generative AI lunch? The networking effect of all devs porting their LLMs etc. to that card would instantly put them as a major CUDA threat. But beancounters running the company would never get such an idea...

This is a gaming card. Look at benchmarks.

Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

#62
post #53

Why don't they just release a basic GPU with 128GB RAM and eat NVidia's local generative AI lunch? The networking effect of all devs porting their LLMs etc. to that card would instantly put them as a major CUDA threat. But beancounters running the company would never get such an idea...

Just how "basic" do you think a GPU can be while having the capability to interface with that much DRAM? Getting there with GDDR6 would require a really wide memory bus even if you could get it to operate with multiple ranks. Getting to 128GB with LPDDR5x would be possible with the 256-bit bus width they used on the top parts of the last generation, but would result in having half the bandwidth of an already mediocre…

M3/M4 Max MacBooks with 128GB RAM are already way better than an A6000 for very large local LLMs. So even if the GPU is as slow as the one in M3/M4 Max (<3070), and using some basic RAM like LPDDR5x it would still be way faster than anything from NVidia.

Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

#63

Why don't they just release a basic GPU with 128GB RAM and eat NVidia's local generative AI lunch? The networking effect of all devs porting their LLMs etc. to that card would instantly put them as a major CUDA threat. But beancounters running the company would never get such an idea...

Meta comment: "why don't they just" phrase usually indicates significant ignorance about a subject, it's better to learn a little bit before dispensing criticism about beancounters or whatnot.

In this case, the die I/O limits precludes more than a reasonable number of DDR channels.

Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

#64

Why don't they just release a basic GPU with 128GB RAM and eat NVidia's local generative AI lunch? The networking effect of all devs porting their LLMs etc. to that card would instantly put them as a major CUDA threat. But beancounters running the company would never get such an idea...

GDDR isnt like the ram that connects to cpu, it's much more difficult and expensive to add more. You can get up to 48GB with some expensive stacked gddr, but if you wanted to add more stacks you'd need to solve some serious signal timing related headaches that most users wouldn't benefit from. I think the high memory local inference stuff is going to come from "AI enabled" cpus that share the memory in your computer.…

They can use LPDDR5x, it would still massively accelerate inference of large local LLMs that need more than 48GB RAM. Any tensor swapping between CPU RAM and GPU RAM kills the performance.

Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

#65
post #51

Earlier quoted context omitted.

48 GB is at the tail end of what's reasonale for normal GPUs. The IO requires a lot of die space. And intel's architecture is not very space efficient right now compared to nvidia's

> The IO requires a lot of die space. And even if you spend a lot of die space on memory controllers, you can only fit so many GDDR chips around the GPU core while maintaining signal integrity. HBM sidesteps that issue but it's still too expensive for anything but the highest end accelerators, and the ordinary LPDDR that Apple uses is lacking in bandwidth compared to GDDR, so they have to compensate with ginormous am…

Going off of how the 4090 and 7900 xtx is arranged I think you could maybe fit on or two chips more around the die over their 12, but that's still a far cry from 128. That would probably just need a shared bus like normal DDR as you're not fitting that much with 16 gbit density

Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

#66

Earlier quoted context omitted.

Judging by the number of 16 GB laptops I see around, 128 GB of RAM would probably cost a bajillion dollars

[flagged]

They all look like amusement parks nowadays.

Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

#68

Earlier quoted context omitted.

Because if they could just do that and it would rival what NVidia has, they would just do it. But obvoiusly they don't. And for reasons: NVidia has worked on CUDA for ages, do you believe they just replace this whole thing in no time?

llama.cpp and its derivatives say yes.

A fraction of CUDA capabilities.

Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

#69

I wanted to have alternative choices than Nvidia for high power GPUs. Then the more I thought about it, the more it made sense to rent cloud services for AI/ML workloads and lesser powered ones for gaming. The only use cases I could come up with for wanting high-end cards are 4k gaming (a luxury I can't justify for infrequent use) or for PC VR which may still be valid if/when a decent OLED (or mini-OLED) headset is a…

Don't rent a GPU for gaming, unless you're doing something like a full-on game streaming service. +10ms isn't much for some games, but would be noticeable on plenty.

IMO you want those frames getting rendered as close to the monitor as possible, and you'd probably have a better time with lower fidelity graphics rendered locally. You'd also get to keep gaming during a network outage.

Re: Intel announces Arc B-series "Battlemage" discrete graphics with Linux support

#70
post #56

Why don't they just release a basic GPU with 128GB RAM and eat NVidia's local generative AI lunch? The networking effect of all devs porting their LLMs etc. to that card would instantly put them as a major CUDA threat. But beancounters running the company would never get such an idea...

They are probably held back by same reason thats preventing AMD and nVidia from doing it either.

NVidia and AMD make $$$ on datacenter GPUs so it makes sense they don't want to discount their own high-end. Intel has nothing there so they can happily go for commodization of AI hardware like what Meta did when releasing LLaMA to the wild.
Post reply on HN