Live data from Hacker News

Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

cnbc.com

131–140 of 340 posts

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#131
Like a lot of the commenters here, I have a problem with this headline. They don't "seek to become a platform company", as in making their own cloud platform where they rent GPU time, meaning they stop selling GPUs to other cloud platforms. That easy misinterpretation makes good clickbait, but no, that's not what the article says - the article has Huang bragging that CUDA already is a parallel computing platform, for a decade or more, and Blackwell Architecture is so integrated and customizable with CUDA (with all its user-extendable kernels and community) that it's thought of as a platform rather than just chip architecture.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#132
post #131

Like a lot of the commenters here, I have a problem with this headline. They don't "seek to become a platform company", as in making their own cloud platform where they rent GPU time, meaning they stop selling GPUs to other cloud platforms. That easy misinterpretation makes good clickbait, but no, that's not what the article says - the article has Huang bragging that CUDA already is a parallel computing platform, for…

[deleted]

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#134
post #105

Platform co seems fitting, considering Nvidia's data center revenue in the fourth quarter of 2023 was a record $18.4 billion, which is 27% higher than the previous quarter and 409% higher than the previous year. Seems revenue from inference is growing at a significant clip.

[deleted]

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#135
post #64

What is FP4, 4 bit floating point? If so, the comparison graph [0] with 30x above Hopper was a bit misleading. [0] https://youtu.be/Y2F8yisiS6E?t=4698

It's 4 bit floating point, at twice the speed of 8 bit floating point. There's also FP6, doesn't offer faster compute than FP8 but manages to take advantage of the better memory bandwith and cache use of the 6 bit format.

Apparently some people are drawing connections to this paper [1] on 4 bit LLMs, which has one NVIDIA employee among its contributors

1: https://arxiv.org/pdf/2310.16836.pdf

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#136
post #34

Earlier quoted context omitted.

Because their stock value is highly coupled with crypto mining and AI craze. The move from PoW to PoS for most crypto networks in combination with bust of ‘22. NVDA slid down in value. OpenAI debuts ChatGPT in late 2022 and now it’s suddenly bumping in price as the hype and rush for GPUs from companies of all types buys up their stock of GPUs. Demand is far outpacing the supply. Nvda can’t keep up. Thus, share price…

If you are a true believer that AI is not a craze, then the stock can only go up from here. If you think there is a chance that everyone gets bored of AI and moves on to some other fad that is not in Nvidia’s wheelhouse, then it’s probably down from here. I’m staying out of this bet: don’t have the stomach for it.

> If you think there is a chance that everyone gets bored of AI and moves on to some other fad that is not in Nvidia’s wheelhouse, then it’s probably down from here.

You may wish to look at history to see how things can work out: Cisco had a P/E ratio of 148 in 1999:

* https://www.dividendgrowthinvestor.com/2022/09/cisco-systems...

The share price tanked, but that does not mean that people got bored of the Internet and the need for routers and switches. QCOM had a P/E of 166: did people decide that mobile communications was a fad?

The connection between technological revolutions and financial bubbles dates back to (at least) Canal Mania:

* https://en.wikipedia.org/wiki/Canal_Mania

* https://en.wikipedia.org/wiki/Technological_Revolutions_and_...

It is possible for both AI to be a big thing and for NVDA to drop.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#137
post #121
post #64

What is FP4, 4 bit floating point? If so, the comparison graph [0] with 30x above Hopper was a bit misleading. [0] https://youtu.be/Y2F8yisiS6E?t=4698

How can 4 bits possibly be enough? Are intermediate calculations done at a higher width and then converted down back to FP4?

For training FP4 sounds pretty niche, but for inference it might be very useful.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#138
post #121
post #64

What is FP4, 4 bit floating point? If so, the comparison graph [0] with 30x above Hopper was a bit misleading. [0] https://youtu.be/Y2F8yisiS6E?t=4698

How can 4 bits possibly be enough? Are intermediate calculations done at a higher width and then converted down back to FP4?

- Training isn’t done at 4-bits, to date this small size has only been for inference.

- Research for a while now has been finding that smaller weights are surprisingly effective. It’s kind of a counterintuitive result, but one way to think about it is there are billions of weights working together. So taken as a whole you still have a large amount of information.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#139
post #88
post #77

Earlier quoted context omitted.

Probably better to stick to the GPUs. Integration is a low margin game.

I'd prefer they stick to GPUs, but I think you're over simplifying. Dell proves that selling complete units is very profitable. Apple shows that owning the entire stack is immensely profitable. Nvidia already has significant hardware and software investment. They very well could fully integrate and grab larger slices of the pie. In fact, Nvidia already has complete appliance like fully integrated machines. But enterp…

>> Apple shows that owning the entire stack is immensely profitable.

Apple shows no such thing. Apple, sells pretty, reliable and safe. A car is a car, but apple is a sports car, or a saloon. Vertical integration is the way they chose to deliver that, and pretty and reliable are all normal people care about.

Nvidia is gonna have to think long and hard about the "whole stack". 20 years ago they might have been able to pull a next, but right now anything that isnt LINUX is a rounding error, and they dont want to turn into SUN (no one at nividia is smart enough to make them the next sun).

Nvidia architecture + Nvidia os is not something that I see them pulling off for the datacenter.

Re: Nvidia CEO Jensen Huang announces new AI chips: ‘We need bigger GPUs’

#140
post #87
post #66

I think at this point, they should stop making it video “cards” but rather video “stations”, a full tower station with power supply and one giant “card” inside with proper cooling, etc., might also justify the crazy prices anyway.

https://lambdalabs.com/gpu-workstations/vector

Heh Heh Heh

"Ultimate AI Workstation". Pricing starts at US$83,549:

https://shop.lambdalabs.com/gpu-workstations/vector/customiz...

Adding every option only adds $2,100 to the price too (totalling $85,649). They should probably just include everything as standard. ;)

Post reply on HN