Earlier quoted context omitted.
I think it will be a winning strategy. Lag is a real killer for LLMs. I think they'll have another LLM on a server (maybe a deal for openai/gemini) that the one on the device can use like ChatGPT uses plugins. But on device Apple have a gigantic advantage. Rabbit and Humane are good ideas humbled by shitty hardware that runs out of battery, gets too hot, has to connect to the internet to do literally anything. Apple…
There really isn't enough emphasis on the downsides of server side platforms. So many of these are only deployed in US and so if you're say in country Australia not only do you have all your traffic going to the US but it will be via slow and intermittent cellular connections. It makes using services like LLMs unusably slow. I miss the 90s and having applications and data reside locally.
Generally an LLM seems to take about 3s or more to respond, and the network delay to the US is a couple of hundred milliseconds.
The network delay seems minimal compared to the actual delay of the LLM.