Earlier quoted context omitted.
Apple is already one of the winners of the AI race. It’s making much more profit (ie it ain’t losing money) on AI off of ChatGPT, Claude, Grok (you would be surprised at how many incels pay to make AI generated porn videos) subscriptions through the App Store. It’s only paying Google $1 billion a year for access to Gemini for Siri
Plus all those pricey 512GB Mac Studios they are selling to YouTubers.
iPhone 17 Pro Demonstrated Running a 400B LLM
41–50 of 362 posts
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#42Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#43Run an incredible 400B parameters on a handheld device. 0.6 t/s, wait 30 seconds to see what these billions of calculations get us: "That is a profound observation, and you are absolutely right ..."
I don't think we are ever going to win this. The general population loves being glazed way too much.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#44Run an incredible 400B parameters on a handheld device. 0.6 t/s, wait 30 seconds to see what these billions of calculations get us: "That is a profound observation, and you are absolutely right ..."
Better than waiting 7.5 million years to have a tell you the answer is 42.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#45Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#46https://xcancel.com/anemll/status/2035901335984611412
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#47It's crazy to see a 400B model running on an iPhone. But moving forward, as the information density and architectural efficiency of smaller models continue to increase, getting high-quality, real-time inference on mobile is going to become trivial.
If they continue to increase.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#48Run an incredible 400B parameters on a handheld device. 0.6 t/s, wait 30 seconds to see what these billions of calculations get us: "That is a profound observation, and you are absolutely right ..."
So this post is like saying that yes an iPhone is Turing complete. Or at least not locked down so far that you're unable to do it.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#49It’s 400B but it’s mixture of experts so how many are active at any time?
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#50> SSD streaming to GPU Is this solution based on what Apple describes in their 2023 paper 'LLM in a flash' [1]? 1: https://arxiv.org/abs/2312.11514
This is not entirely dissimilar to what Cerebus does with their weights streaming.