Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
1–10 of 196 posts
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#2is there a comparison of it running on iPhone vs. Android phones?
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#3It runs on Android too, with AI Core or even with llama.cpp
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#4Is the output coherent though? I am yet to see a local model working on consumer grade hardware being actually useful.
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#5Is the output coherent though? I am yet to see a local model working on consumer grade hardware being actually useful.
I run qwen3.5 122b on a Framework Desktop at 35/ts as a daily driver doing security and OS systems and software engineering.
Never paid an LLM provider and I have no reason to ever start.
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#6Is the output coherent though? I am yet to see a local model working on consumer grade hardware being actually useful.
I can try it for you
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#7is there a comparison of it running on iPhone vs. Android phones?
You can run Android on just about anything so it boils down to Linux GPU benchmarks.
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#8Is the output coherent though? I am yet to see a local model working on consumer grade hardware being actually useful.
It can write (some) code that works. Just roughly guessing from my use, but I think of it as being a bit like ChatGPT circa-2024 in terms of capability & speed.
Disappointing if you compare it to anything else from 2026, but fairly impressive for something that can run locally at an OK speed.
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#9[flagged]
Re: Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
#10> edge AI deployment
Isn't the "edge" meant to be computing near the user, but not on their devices?