Live data from Hacker News

Hands-On with the AMD Ryzen AI Halo

microcenter.com

41–45 of 45 posts

Re: Hands-On with the AMD Ryzen AI Halo

#41

It seems like there is a very health space for an MOE targeted GPU where it has essentially an 5070ti with 16gb ish GDDR7 but then also has 128 GB LPDDR5x (or even just DDR5 as expansion dimms on it?). Putting this into the same card would likely reduce the transfer hit when a cache miss happened and the gpu had to load from the slower LPDDR5x. No need to have PCIe 5x16 limiting memory transfer if it is on the card.…

Sounds like Bolt's GPU

their main mem is LPDDR5x with DDR5 SODIMMs for expansion. LPDDR5x is 256GB/s? on the machines you see it implemented on which is a lot faster than the DDR5 expansion. For comparison, gddr7 on a 5070ti pumps ~900GB/s. 4x the tokens/sec if you are memory bound.

Re: Hands-On with the AMD Ryzen AI Halo

#45
amd ai apus are incapable of running models fast enough for productive work, you need over 100t/s for text mode and above 180t/s for image gen/ image recognition to not wait minutes. you cannot run claw and leave it - you would need to wait multile hours for small project like chatbot+landing page. price wise hw was not worth it even when it cost under 2k$, nowadays it cost even more.

just a perspective(you can re check results in youtube reviews): ryzen ai runs most moe models under 60t/s while nvidia gpu can runt them at 100t/s. and you preferably need deep models for code, they would run at 40t/s.

Post reply on HN