Viewing profile — darkolorin
darkolorin
HN member- Joined
- Tue, Aug 05, 2014, 2:15 PM UTC
- HN karma
- 73
- Public activity
- 18 items
- HN profile
- View on Hacker News ↗
About darkolorin
Recent public activity
-
story
Show HN: We made LM studio alternative based on own engine
Hi guys. We made own Mac app for M series hardware. Based on own engine called uzu (available with MIT on GitHub). Completely written from scratch and inference too. Our goal is to…
-
comment
Comment #44575711
Basically “faster” means better performance e.g. tokens/s without loosing quality (benchmarks scores for models). So when we say faster we provide more tokens per second than llama…
-
story
Show HN: We made our own inference engine for Apple Silicon
We wrote our inference engine on Rust, it is faster than llama cpp in all of the use cases. Your feedback is very welcomed. Written from scratch with idea that you can add support …
- comment
- story
- comment
- comment
- story
-
comment
Comment #43553251
I made it! 90 t/s on my iPhone with llama1b fp16 We completely rewrite the inference engine and did some tricks. This is a summarization with llama 3.2 1b float16. So most of the t…
- story
- story
- story
- story
- story
- story
-
comment
Comment #8669306
It's my first experience on Medium. I hope community can help me to improve my skills.
- story
-
comment
Comment #8236777
really great idea, but why morning? %)