Earlier quoted context omitted.
I just tested Qwen3 A3B vs ChatGPE a random prompt from my head and: > Please write a C# middleware to block requests from browser agents that contain any word in a specified list of words: openai, grok, gemini, claude. I used ChatpGPT 4o from GitHub Copilot inside VSCode. And Qwen3 A3B from here: https://deepinfra.com/Qwen/Qwen3-30B-A3B ChatGPT 4o was considerably better. Less verbose and less unnecessary abstractio…
You want the 2507 update of the model, I think the one you used is ~8-10 months out-of-date.
Experimenting with Local LLMs on macOS
161–170 of 276 posts
Re: Experimenting with Local LLMs on macOS
#162So far I've not run into the kind of use cases that local LLMs can convincingly provide without making me feel like I'm using the first ever ChatGPT from 2022, in that they are limited and quite limiting. I am curious about what use cases the community has found that work for them. The example that one user has given in this thread about their local LLM inventing a Sun Tzu interview is exactly the kind of limitation…
Re: Experimenting with Local LLMs on macOS
#163Earlier quoted context omitted.
I feel like Apple needs a new CEO, I've felt this way for a long time. If I had been in charge of Apple I would have embraced local LLMs and built an inference engine that optimizes models that are designed for Nvidia, I also would have probably toyed around with the idea of selling server-grade Apple Silicon processors and opening up the GPU spec so people can build against it. Seems like Apple tries to play it too…
I think shareholders are fine with Tim Cook as a CEO.
Its easy to sit in the armchair and say "just be a visionary bro" when they forget Tim worked under Steve for awhile before his death - he has some sense and understanding of what it takes to get a great product out of the door.
Nvidia is generating a lot of revenue, sure - but what is the downstream impact on its customers with the hardware? All they have right now is negative returns to show for their spending. Could this change? Maybe. Is it likely? Not in my view.
As it stands, Apple has made the absolute right choice in not wasting its cash and is demonstrating discipline. Which when all this LLM mania quietens, shareholders will respect.
Re: Experimenting with Local LLMs on macOS
#164Earlier quoted context omitted.
I feel like Apple needs a new CEO, I've felt this way for a long time. If I had been in charge of Apple I would have embraced local LLMs and built an inference engine that optimizes models that are designed for Nvidia, I also would have probably toyed around with the idea of selling server-grade Apple Silicon processors and opening up the GPU spec so people can build against it. Seems like Apple tries to play it too…
I think if Cook had vision, he could have started something called Apple Enterprise and sold Apple Silicon as a server and made AI chips. I agree he’s too conservative and has no product vision. Great manager though.
Re: Experimenting with Local LLMs on macOS
#165Earlier quoted context omitted.
I think if Cook had vision, he could have started something called Apple Enterprise and sold Apple Silicon as a server and made AI chips. I agree he’s too conservative and has no product vision. Great manager though.
[flagged]
Re: Experimenting with Local LLMs on macOS
#166Earlier quoted context omitted.
I thought Apple MLX can do that if you convert your model using it https://mlx-framework.org/
MLX does not support the ANE. https://github.com/ml-explore/mlx/issues/18
That’s just an issue with stale and incorrect information. Here are the docs https://opensource.apple.com/projects/mlx/
Re: Experimenting with Local LLMs on macOS
#167Check out Osaurus - MIT Licensed, native, Apple Silicon–only local LLM server - https://github.com/dinoki-ai/osaurus
Re: Experimenting with Local LLMs on macOS
#168I don't think we're anywhere close to running cutting-edge LLMs on our phones or laptops. What may be around the corner is running great models on a box at home. The AI lives at home. Your thin client talks to it, maybe runs a smaller AI on device to balance latency and quality. (This would be a natural extension for Apple to go into with its Mac Pro line. $10 to 20k for a home LLM device isn't ridiculous.)
This is what I’m doing with my amd 395+. I’m running docker containers with different apps and it works well enough for a lot of my use cases. I mostly use Qwen Code and GPT OSS 120b right now. When the next generation of this tech comes through I will probably upgrade despite the price, the value is worth it to me.
Re: Experimenting with Local LLMs on macOS
#169Earlier quoted context omitted.
I think if Cook had vision, he could have started something called Apple Enterprise and sold Apple Silicon as a server and made AI chips. I agree he’s too conservative and has no product vision. Great manager though.
[flagged]
Re: Experimenting with Local LLMs on macOS
#170Earlier quoted context omitted.
> But whenever you mention crypto mining or AI datacenter markets, people act like Apple is above selling products that people want. People also want comfortable mattresses and high quality coffee machines. Should Apple make them too? Apple not being in a particular industry is a perfectly valid choice, which is not remotely comparable to protecting their interests in the industries they are currently in. Selling dat…
Apple is perfectly well equipped to sell datacenter products. They've done it in the past, even supporting Nvidia's compute drivers along the way. If they have the staff to design consumer-facing and developer-facing experiences, why wouldn't they address the datacenter? Money is money. 10 years ago people would have laughed at the notion of Nvidia abandoning the gaming market, now it's their most lucrative option. A…