AMD has better seemingly better hardware - but not the production capacity to compete with Nvidia yet. Will be interesting to see margins compress when real competition catches up. Everybody thinks it’s CUDA that makes Nvidia the dominant player. It’s not - almost 40% of their revenue this year comes from mega corporations that use their own custom stack to interact with GPUs. It’s only a matter of time before compet…
Can you explain the cuda-less stack a little more or provide a source?
AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
121–130 of 273 posts
Re: AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
#122"TensorWave is a cloud provider specializing in AI workloads. Their platform leverages AMD’s Instinct™ MI300X accelerators, designed to deliver high performance for generative AI workloads and HPC applications." I suggest taking the report with a grain of salt.
The salt is in the plain sight. The do the standard AMD comparison: 8x AMD MI300X (192GB, 750W) GPU 8x H100 SXM5 (80GB, 700W) GPU The fair comparison would be against 8x H100 NVL (188GB, Price tells a story. If AMD performance would be in par with Nvidia they would not sell their cards for 1/4 price.
Re: AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
#123Earlier quoted context omitted.
[flagged]
[flagged]
>You need big companies with scale that pay huge wages as those mean a lot more tax revenue.
Now provide proofs that we need some big corps dodging taxes, including FAANG, and not more small and middle-sized business.
Re: AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
#124Earlier quoted context omitted.
Why is producing big companies a goal? High standards of living for all seems to a much better goal. And that can be done with small or big companies - so long as economic production is high enough and distributed well enough.
Because tech innovation requires tons of R&D and you can't afford to do that otherwise. Europeans use American laptops running an American operating system to watch American movies in an American browser. European economic production is nowhere near high enough and now Europe is struggling to provide for its aging population and doesn't have enough good jobs for younger people. I support redistribution generally, but…
Re: AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
#125Earlier quoted context omitted.
Same thing was said about Nvidia's crypto bubbles, and then look what happened. Jensen isn't stupid. He's making accelerators for anything so that they'll be ready to catch the next bubble that depends on crazy compute power that can't be done efficiently on CPUs. They're so far the only semi company beating Moore's law by a large margin due to their clever scaling tech while everyone else is like "hey look our new p…
I think there's also a very high prospect of virtual worlds with virtual people (SFW or otherwise) becoming popular, rendered with Apple/META goggles...that could require insane amounts of compute. And this is just one possibility. Relatively cheap multimodal smart glasses you wear when out and around that offload compute to the cloud are another. Nvidia could just as easily triple in short order as get cut in half f…
[0]https://www.macrumors.com/2024/04/23/apple-cuts-vision-pro-s... [1] https://www.macrumors.com/2024/04/22/apple-vision-pro-custom...
Re: AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
#126Are these fabbed at the same process node? (Otherwise it's apples and oranges)
It's not apples and oranges. These are the top of the line offerings from the respective companies today.
Re: AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
#127Earlier quoted context omitted.
> There is one exception, founded in 1984 Europe clearly has many problems stimulating investments and creating a competitive environment for startups and tech companies. But you can't say ASML is the only one "real" tech company. What about Adyen, Spotify, Klarna, N26, Revolut, etc?
Most of those companies you mentioned aren't anywhere near as wealthy or as high market caps as US big-tech. Most of them are just payment middlemen not some innovative product nobody else can do, and Spotify survives on monopolizing and squeezing artists, not some innovative product. Kind of like Netflix except Netflix has some cutting edge streaming tech as a product not just IP licenses. ASML is the only product i…
Re: AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
#128Earlier quoted context omitted.
I would argue that the EU is quite a successful organization given the task of setting up general market rules for 26 countries. Obviously there is a lot to criticize about the EU and I can offer you a gigantic list there too. However, I do not see any clear failure of the EU’s approach as a single market so far. Additionally part of the philosophy was establishing peace in a region that was torn up by wars for a lot…
It's definitely successful, and I was probably too harsh there. But I genuinely think the barriers that are left are damn near insurmountable. An awful lot has to change before a Greek tech workers can move to Sweden as easily as a Virginian can move to California.
Re: AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
#129Earlier quoted context omitted.
>How can a wealthy continent with 750 million people produce no big tech companies? It's a big problem. Much more difficult to scale a product across 26 different countries and nearly as many languages and regulatory jurisdictions. US is one country, not a collection of countries fighting each other, meaning your product is instantly available to 300M people speaking the same language under (nearly) the same regulati…
>US is one country, not a collection of countries fighting each other The US is a republic of 50 states. Each state has a huge amount of sovereignty and autonomy. There are 50 state-level regulatory jurisdictions. Not to mention the local-level of government. But in spite of this, the US does not over-regulate. This is the big difference to Europe (I say this as an American expat living in Europe).
Re: AMD's MI300X Outperforms Nvidia's H100 for LLM Inference
#130Why the hell are we doing 128 input token benchmarks in 2024. This is not representative of most workloads, and prefill perf is incredibly important.
For understanding: What would be a suitable input length in your oppinion? And why isnt this a good one: Are real-life queries shorter? Or longer? If i count one word as a token, then in my case most of the queries are less than 128 words.