Earlier quoted context omitted.
I don't think most dev laptops have 8GB+ of GPU memory, which (based on ollama requirements) seems like it's on the low-mid end of the requirements. I've tried experimenting with some local models and: a) They're much slower on my 6GB laptop GPU b) The seem to not be as good, functionally b) I can't make use of larger models I haven't done more than just some experimentation but I can see how this would make sense to…
MacBooks especially pros are in a very privileged position due to their memory architecture. With some dedicated hardware it may be reasonable to run, perhaps even train, useful LLMs on device. I guess we’ll see in a year or two. This must be on everyone’s radar now, Apple won’t be the odd man out.
Lisa Su was interviewed at Code recently, and the AI discussion focused on competing against NVIDIA in big datacenters, but I hope they're also thinking about client-side stuff like that.