I just open-sourced how https://github.com/mylovelycodes/LiteRTLM-Swift

LiteRTLM-Swift lets you run LLMs locally with a clean Swift API. - On-device inference - No cloud required - Built for iOS

Show HN: Running Gemma 4 on an iPhone 13 Pro
github.com