It will be funny if we go back to lugging around brick-size batteries with us everywhere!
iPhone 17 Pro Demonstrated Running a 400B LLM
121–130 of 362 posts
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#122Earlier quoted context omitted.
Only way to have hardware reach this sort of efficiency is to embed the model in hardware. This exists[0], but the chip in question is physically large and won't fit on a phone. [0] https://www.anuragk.com/blog/posts/Taalas.html
That's actually pretty cool, but I'd hate to freeze a models weights into silicon without having an incredibly specific and broad usecase.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#123Earlier quoted context omitted.
Looked at a certain way it's incredible that a 40-odd year old comedy sci-fi series is so accurate about the expected quality of (at least some) AI output. Which makes it even funnier. It makes me a little sad that Douglas Adams didn't live to see it.
Also check out "The Great Automatic Grammatizator" by Roald Dahl for another eerily accurate scifi description of LLMs written in 1954: https://gwern.net/doc/fiction/science-fiction/1953-dahl-theg...
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#124I have some macro opinions about Apple - not sure if I'm correct, but tell me what you think. Apple has always seen RAM as an economic advantage for their platform: Make the development effort to ensure that the OS and apps work well with minimal memory and save billions every year in hardware costs. In 2026, iPhones still come with 8Gb of RAM, Pro/Max come with 12Gb. The problem is that AI (ML/LLM training and infer…
Why do you say they can't do this?
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#125Earlier quoted context omitted.
Only way to have hardware reach this sort of efficiency is to embed the model in hardware. This exists[0], but the chip in question is physically large and won't fit on a phone. [0] https://www.anuragk.com/blog/posts/Taalas.html
That's actually pretty cool, but I'd hate to freeze a models weights into silicon without having an incredibly specific and broad usecase.
The $$$ would probably make my eyes bleed tho.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#126It's crazy to see a 400B model running on an iPhone. But moving forward, as the information density and architectural efficiency of smaller models continue to increase, getting high-quality, real-time inference on mobile is going to become trivial.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#127A year ago this would have been considered impossible. The hardware is moving faster than anyone's software assumptions.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#128It's crazy to see a 400B model running on an iPhone. But moving forward, as the information density and architectural efficiency of smaller models continue to increase, getting high-quality, real-time inference on mobile is going to become trivial.
> moving forward, as the information density and architectural efficiency of smaller models continue to increase If they continue to increase.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#129It will be funny if we go back to lugging around brick-size batteries with us everywhere!