Earlier quoted context omitted.
> Apple is about the only company that allows you to run any decently sized LLMs on a laptop Comments like these make me wonder if people even use different computers anymore or if they just say stuff and hope that it's right.
Please, do tell me where else can I run 32b models locally without quantization
On any other computers with integrated GPU you can have more memory than on Apple computers, where memory is abnormally expensive.
The main advantage of the current Apple computers is a wider interface with memory, which provides a better memory throughput for their integrated GPUs.
In 2025 it is expected that others will catch up with Apple even from this point of view (e.g. AMD Strix Halo).