Live data from Hacker News

Meta MTIA v2 – Meta Training and Inference Accelerator

ai.meta.com

31–40 of 63 posts

Re: Meta MTIA v2 – Meta Training and Inference Accelerator

#31
post #20

I find it weird that not everyone agree Meta and Facebook and social networks in general are doing some good the the society and our democracies; yet they manage to spend incredible amount of money/energy/time to develop solutions to problems we aren't exactly sure are worth solving…

What is worth solving in your opinion? Should they not make their service more efficient?

I assume this helps reduce their server and electricity costs. At a certain scale these things pay off.

Re: Meta MTIA v2 – Meta Training and Inference Accelerator

#32
post #2

Intel Gaudi 3 has more interconnect bandwidth than this has memory bandwidth. By a lot. I guess they can't be fairly compared without knowing the TCO for each. I know in the past Google's TPU per-chip specs lagged Nvidia but the much lower TCO made them a slam dunk for Google's inference workloads. But this seems pretty far behind the state of the art. No FP8 either.

Also its at 90 watts vs 900 watts for gaudi 3, the flops/mem bw per watt is much more comparable.

Re: Meta MTIA v2 – Meta Training and Inference Accelerator

#33
post #2

Intel Gaudi 3 has more interconnect bandwidth than this has memory bandwidth. By a lot. I guess they can't be fairly compared without knowing the TCO for each. I know in the past Google's TPU per-chip specs lagged Nvidia but the much lower TCO made them a slam dunk for Google's inference workloads. But this seems pretty far behind the state of the art. No FP8 either.

Also its at 90 watts vs 900 watts for gaudi 3, the flops/mem bw per watt is much more comparable.

It would be interesting if this could be made into a reasonably priced (lmao) card for home inference if they intend to mass produce it.

Can't imagine any other reason other than cost as to why they went with LPDDR5, LPDDR5X has more bandwidth and GDDR6 has even more.

Re: Meta MTIA v2 – Meta Training and Inference Accelerator

#35

I thought MTIA v2 would use the mx formats https://arxiv.org/pdf/2302.08007.pdf , guess they were too far along in the process to get it in this time. Still this looks like it would make for an amazing prosumer home ai setup. Could probably fit 12 accelerators on a wall outlet with change for a cpu, would have enough memory to serve a 2T model at 4bit and reasonable dense performance for small training runs and image…

I suppose it might, there are not a lot of details (what kind of sparsity for example?) about what they mean in terms of INT8 support - it could be MXINT8, or something else. Glad someone was thinking the same thing I was though!

its gotta be that 2/4 sparsity that everyone has, but I haven't seen used anywhere right? If they put it in though they must be using it, but I'm not sure for what. And without details I think its a good bet that int8 is the standard int8.

Wishful thinking maybe they'll announce selling it with the giant llama3 cause there's no good, cheap way to inference something like that at home at the moment and this could change that.

Re: Meta MTIA v2 – Meta Training and Inference Accelerator

#36

Earlier quoted context omitted.

Also its at 90 watts vs 900 watts for gaudi 3, the flops/mem bw per watt is much more comparable.

It would be interesting if this could be made into a reasonably priced (lmao) card for home inference if they intend to mass produce it. Can't imagine any other reason other than cost as to why they went with LPDDR5, LPDDR5X has more bandwidth and GDDR6 has even more.

they didn't use GDDR cause they wanted the memory capacity which is really important for recommendation models. But I totally agree that this is a sort of perfect cost/perf per watt point for a home setup. I really hope they do it, if not for this one at least for v3.

Re: Meta MTIA v2 – Meta Training and Inference Accelerator

#37
post #24

Earlier quoted context omitted.

To me, it’s bizarre to see the HPC mindset taking hold again after the cloud/commodity mindset dominated the last 16 years. You don’t always need a Ferrari to go to the store

WDYM by HPC mindset?

"The only meaningful benchmark in the world is LAPACK and only larger than ever monolithic problem instances matter, I don't know what you're talking about, 'embarrassingly parallel'? What a silly word! Serving web requests concurrently? Good for you, congratulations, but can you do parallel programming?"

Sorry if this make anyone feels bad. It certainly made myself uncomfortable typing it out though.

Re: Meta MTIA v2 – Meta Training and Inference Accelerator

#40
post #13

I like the interactive 3D widget showing off the chip. Yep, that sure is a metal rectangle.

Really annoys me that the loading animation of these before-/after-images doesn't finish on firefox and that it won't let me drag the knob with the separator. ...no "Under the hood" for me.

Dragging the top-left corner works, for some reason. Really bizarre UI issue.
Post reply on HN