Hi guys. We made own Mac app for M series hardware. Based on own engine called uzu (available with MIT on GitHub).
Completely written from scratch and inference too. Our goal is to show what you can do with local models. In the app section we will publish use cases which you can easily implement in your app.
Would appreciate any feedback.
Right now there is no set of proper benchmarks from us. In general we are faster than llama.cpp and on par with MLX. In a next release we will be faster than both.
Show HN: We made LM studio alternative based on own engine
trymirai.com