Unfortunately Apple appears to be blocking the use of these llms within apps on their app store. I've been trying to ship an app that contains local llms and have hit a brick wall with issue 2.5.2
Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
31–40 of 196 posts
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#32Unfortunately Apple appears to be blocking the use of these llms within apps on their app store. I've been trying to ship an app that contains local llms and have hit a brick wall with issue 2.5.2
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#33Unfortunately Apple appears to be blocking the use of these llms within apps on their app store. I've been trying to ship an app that contains local llms and have hit a brick wall with issue 2.5.2
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#34Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#35Earlier quoted context omitted.
I run qwen3.5 122b on a Framework Desktop at 35/ts as a daily driver doing security and OS systems and software engineering. Never paid an LLM provider and I have no reason to ever start.
What spec of Framework Desktop do you run this on?
The only downside is that I suspect the Framework would be a decent bit quieter under load (not that this thing is abnormally loud). As well as you're limited to a single M.2 2230 internal SSD slot in this (I believe Micron recently launched a 4 TB model, but generally you'll max out at 2 TB without using an external enclosure).
I don't have anything against the Framework, I'm sure it's a great machine, but the Z13 is an incredible portable all-in-one device that can handle everything from general PC use to gaming to tablet/entertainment to LLMs & high perf.
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#36Unfortunately Apple appears to be blocking the use of these llms within apps on their app store. I've been trying to ship an app that contains local llms and have hit a brick wall with issue 2.5.2
Though of course Apple's rules aren't always consistent, I have 2 separate apps currently on my phone that can/are running this (Google's Edge Gallery and Locally AI)
But it's more likely it's just walled garden + security theatre that'll keep them from allowing outside apps.
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#37For those who would like an example of its output, I'm currently working through creating a small, free (cc0, public domain) encyclopedia (just a couple of thousand entries) of core concepts in Biology and Health Sciences, Physical Sciences, and Technology. Each entry is being entirely written by Gemma 4:e4b (the 10 GB model.) I believe that this may be slightly larger than the size of the model that runs locally on…
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#38Is the output coherent though? I am yet to see a local model working on consumer grade hardware being actually useful.
It's a 100% replacement for free ChatGPT/Gemini.
Compared to the paid pro/thinking models... Gemma does have reasoning, and I have used the reasoning mode for some tax & legal/accounting advice recently as well as other misc problems. It's worked well for that, but I haven't tried any real difficult tasks. From what I've heard re. agentic coding, the open weight models are ~18-24 months behind Anthropic & Google's SOTA.
Qwen 3.5 122B-A10B should just fit into 128 GB with a Q4/5 and may be a bit smarter. There's apparently also a similar sized Gemma 4 model but they haven't released it yet, the 26B was the largest released.
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#39Unfortunately Apple appears to be blocking the use of these llms within apps on their app store. I've been trying to ship an app that contains local llms and have hit a brick wall with issue 2.5.2
I think Apple will become increasingly draconian about LLMs. Very soon people won't need to buy many of their apps. They can just make them. This threatens Apple's entire business model.
A kid playing Roblox can spend more than that in a good weekend.
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#40is there a comparison of it running on iPhone vs. Android phones?
The model itself works absolutely fine, though the iPhone thermal throttles at some point which really reduces the token generation speed. When I asked it to write me a business plan for a fish farm in the Nevada desert, it slowed down after a couple thousand tokens, whereas the Pixel seems to just keep going.