My iPad Air with M2 can run local LLMs rather well. But it gets ridiculously hot within seconds and starts throttling.
I wonder if anyone has made a liquid cooling system for ipads / phones. Like, a sealed thing that seals onto the back of the device and circulates cooling water directly against the back surface.
iPhone 17 Pro Demonstrated Running a 400B LLM
221–230 of 362 posts
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#222This is awesome! How far away are we from a model of this capability level running at 100 t/s? It's unclear to me if we'll see it from miniaturization first or from hardware gains
It will never be possible on a smart phone. I know that sounds cynical, but there's basically no path to making this possible from an engineering perspective.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#223A year ago this would have been considered impossible. The hardware is moving faster than anyone's software assumptions.
Does iPhone have some kind of hardware acceleration for neural netwoeks/ai ?
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#224Earlier quoted context omitted.
The Bistromathics? That's not incorrect, it's simply too advanced for us to understand.
“What do you get if you multiply six by nine?” (One) source: https://www.reddit.com/r/Fedora/comments/1mjudsm/comment/n7d...
To quote the message from the universes creators to its creation “We apologise for the inconvenience”. Does seem to sum up Douglas Adam’s views on absurdity of life.
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#225My iPad Air with M2 can run local LLMs rather well. But it gets ridiculously hot within seconds and starts throttling.
I wonder if anyone has made a liquid cooling system for ipads / phones. Like, a sealed thing that seals onto the back of the device and circulates cooling water directly against the back surface.
https://onexplayerstore.com/products/onexplayer-super-x?vari...
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#226Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#227Earlier quoted context omitted.
I wonder if anyone has made a liquid cooling system for ipads / phones. Like, a sealed thing that seals onto the back of the device and circulates cooling water directly against the back surface.
You can buy a liquid cooled tablet. https://onexplayerstore.com/products/onexplayer-super-x?vari...
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#228It's like the sloth from Zootopia
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#229This is awesome! How far away are we from a model of this capability level running at 100 t/s? It's unclear to me if we'll see it from miniaturization first or from hardware gains
Probably 15 to 20 years, if ever. This phone is only running this model in the technical sense of running, but not in a practical sense. Ignore the 0.4tk/s, that's nothing. What's really makes this example bullshit is the fact that there is no way the phone has a enough ram to hold any reasonable amount of context for that model. Context requirements are not insignificant, and as the context grows, the speed of the o…
Re: iPhone 17 Pro Demonstrated Running a 400B LLM
#230> SSD streaming to GPU Is this solution based on what Apple describes in their 2023 paper 'LLM in a flash' [1]? 1: https://arxiv.org/abs/2312.11514
Yes. I collected some details here: https://simonwillison.net/2026/Mar/18/llm-in-a-flash/