Live data from Hacker News

Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

cnbc.com

71–80 of 340 posts

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#71
post #36

Earlier quoted context omitted.

Tell me more about why you believe their stock is hilariously overvalued.

72 P/E ratio while they have a mere monopoly on one the most valuable resource in the world. Competition WILL come. Maybe it's Groq, maybe AMD, maybe Cerebras. Maybe there's a stealth startup out there. Point is, they're going to be challenged soon.

You and what fab?

It's almost impossible to manufacture at scale with good yields and leading edge fabs are almost all bought out.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#74
post #9
post #6

FP8 being 2.5x Hopper is kind of disappointing after such a long time. Since its 2 fused chips, that means it’s 25% effective delta. though it seems most of the progress has been on memory throughput and power use which is still very impressive. I wonder how this will trickle down to the consumer segment.

Jensen revealed later that the LLM inference is 30x due to architectural improvements, it's massive. I don't know if it's latency or just 2-3x performance boost with 30x more customers served in the same chip. Either way, 30x is massive.

But Blackwell in the graph is FP4 whereas Hopper is FP8.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#75

Earlier quoted context omitted.

Tell me more about why you believe their stock is hilariously overvalued.

No moat. Yes, CUDA, but CUDA is maaaaaybe a few tens of billion USD deep and a few (more) years wide. When the rest of the industry saw compute as a vanity market, that was sufficient. Now, it's a matter of time before margins go to, uhhh, less than 90%. Does that make shorting a good idea? I wouldn't count on it. The market can always remain irrational longer than you can remain solvent.

They also bought infiniband which has played a big role in being the best at clustering, though Google's TPU reconfigurable topology stuff seems really cool too.

Tesla went after them with Dojo and has still ended up splurging on big H100 clusters.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#76

Earlier quoted context omitted.

Tell me more about why you believe their stock is hilariously overvalued.

No moat. Yes, CUDA, but CUDA is maaaaaybe a few tens of billion USD deep and a few (more) years wide. When the rest of the industry saw compute as a vanity market, that was sufficient. Now, it's a matter of time before margins go to, uhhh, less than 90%. Does that make shorting a good idea? I wouldn't count on it. The market can always remain irrational longer than you can remain solvent.

And MS and everyone else have plenty of interest in helping AMD commodify CUDA compatibility.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#77
post #66

I think at this point, they should stop making it video “cards” but rather video “stations”, a full tower station with power supply and one giant “card” inside with proper cooling, etc., might also justify the crazy prices anyway.

Probably better to stick to the GPUs. Integration is a low margin game.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#78
post #67
post #23

Earlier quoted context omitted.

From how I understood it, it means they optimised the entire stack from CUDA to the networking interconnects specifically for data centers, meaning you get 30x more inference per dollar for a datacenter. This is probably not fluff, but it's only relevant for a very very specific use-case, ie enterprises with the money to buy a stack to serve thousands of users with LLMs. It doesn't matter for anyone who's not microso…

They showed 30x was for FP4. Who is using FP4 in practice?

But maybe you should. Once the software stack is ready for it there'll be more people since the performance gains are so massive.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#79
post #60

Earlier quoted context omitted.

How is 2.5x disappointing in one generation?

Compare to the 10x that was Hopper uplift.

Because it involved scaling in chip area needed for FP8. AI community realized that FP8 training is possible few years back so the transistors given for FP8 was scaled. Overall I think transistors grow just by ~50% per generation so most of the gains comes from removing FP32/FP64 share which were dominant 10 years back, but there is only some point it could go to.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#80
post #41

Earlier quoted context omitted.

They are priced as if they are the only ones who are capable of creating chips that can crunch LLM algos. But AMD, Google, Intel, and even Apple are also capable. Apple is in talks with Google to bring Gemini to the iPhone, and it will obviously also be on android phones. So almost every phone on earth is poised to be using Gemini in the near future, and Gemini runs entirely on Google's own custom hardware (which is…

This seems as good a place as any to be Corrected by the Internet, so... correct me if I'm wrong. Making a graphics chip that is as good as Nvidia: Very difficult. Huge moat, huge effort, lots of barriers, lots of APIs, lot of experience, lots of decades of experience to overcome. Making something that can run a NN: Much, much easier. I'd guess, start-up level feasible. The math is much simpler. There's a lot of it,…

CUDA is a big reason for their moat. And that's not something you can build in a couple of years no matter how money you can throw on it.

Without CUDA you have a chip that runs on premise without anyone having a clue how good that is which is supposedly what Google does. Your only offering is cloud services. As big as this is, corporations would want to build their own datacenters.

Post reply on HN