Earlier quoted context omitted.
Maybe you should have asked a better question. :P
What do you get if you multiply six by nine?
iPhone 17 Pro Demonstrated Running a 400B LLM
81–90 of 362 posts
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#82Earlier quoted context omitted.
Possibly this just isn't the generation of hardware to solve this problem in? We're like, what three or four years in at most, and only barely two in towards AI assisted development being practical. I wouldn't want to be the first mover here, and I don't know if it's a good point in history to try and solve the problem. Everything we're doing right now with AI, we will likely not be doing in five years. If I were run…
If I was running a company like Apple, I'd be working with Khronos to kill CUDA since yesterday. There are multiple trillions of dollars that could be Apple's if they sign CUDA drivers on macOS, or create a CUDA-compatible layer. Instead, Apple is spinning their wheels and promoting nothingburger technology like the NPU and MPS. It's not like Apple's GPU designs are world-class anyways, they're basically neck-and-nec…
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#83Earlier quoted context omitted.
I don't think we are ever going to win this. The general population loves being glazed way too much.
The other day, I got: "You are absolutely right to be confused" That was the closest AI has been to calling me "dumb meatbag".
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#84Earlier quoted context omitted.
If I was running a company like Apple, I'd be working with Khronos to kill CUDA since yesterday. There are multiple trillions of dollars that could be Apple's if they sign CUDA drivers on macOS, or create a CUDA-compatible layer. Instead, Apple is spinning their wheels and promoting nothingburger technology like the NPU and MPS. It's not like Apple's GPU designs are world-class anyways, they're basically neck-and-nec…
CUDA is not the real issue, AMD's HIP offers source-level compatibility with CUDA code, and ZLUDA even provides raw binary compatibility. nVidia GPUs really are quite good, and the projected advantages of going multi-vendor just aren't worth the hassle given the amount of architecture-specificity GPUs are going to have.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#85A year ago this would have been considered impossible. The hardware is moving faster than anyone's software assumptions.
This isn't a hardware feat, this is a software triumph. They didn't make special purpose hardware to run a model. They crafted a large model so that it could run on consumer hardware (a phone).
It’s been a lot of years, but all I can hear after reading that is … I’m making a note here, huge success
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#86Run an incredible 400B parameters on a handheld device. 0.6 t/s, wait 30 seconds to see what these billions of calculations get us: "That is a profound observation, and you are absolutely right ..."
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#87It's crazy to see a 400B model running on an iPhone. But moving forward, as the information density and architectural efficiency of smaller models continue to increase, getting high-quality, real-time inference on mobile is going to become trivial.
> moving forward, as the information density and architectural efficiency of smaller models continue to increase If they continue to increase.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#88Earlier quoted context omitted.
Only way to have hardware reach this sort of efficiency is to embed the model in hardware. This exists[0], but the chip in question is physically large and won't fit on a phone. [0] https://www.anuragk.com/blog/posts/Taalas.html
That's actually pretty cool, but I'd hate to freeze a models weights into silicon without having an incredibly specific and broad usecase.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#89Run an incredible 400B parameters on a handheld device. 0.6 t/s, wait 30 seconds to see what these billions of calculations get us: "That is a profound observation, and you are absolutely right ..."
Better than waiting 7.5 million years to have a tell you the answer is 42.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#90Earlier quoted context omitted.
Plus all those pricey 512GB Mac Studios they are selling to YouTubers.
They don't offer the 512 gig RAM variant anymore. Outside of social media influencers and the occasional AI researcher, the market for $10K desktops is vanishingly small.
Pretty sure the M5 Ultra will be out after WWDC, so my M3 Ultra is (while still completely capable of fulfilling my needs) looking a bit long in the tooth. If I can get a good price for it now, I might be able to offset most of the M5 post WWDC...