Live data from Hacker News

Viewing profile — darkolorin

darkolorin

HN member
Joined
Tue, Aug 05, 2014, 2:15 PM UTC
HN karma
73
Public activity
18 items

About darkolorin

prev CEO & co-founder Prisma & Capture, LFG, now doing on-device inference at Mirai

Recent public activity

  1. story
    Show HN: We made LM studio alternative based on own engine

    Hi guys. We made own Mac app for M series hardware. Based on own engine called uzu (available with MIT on GitHub). Completely written from scratch and inference too. Our goal is to…

  2. comment
    Comment #44575711

    Basically “faster” means better performance e.g. tokens/s without loosing quality (benchmarks scores for models). So when we say faster we provide more tokens per second than llama…

  3. story
    Show HN: We made our own inference engine for Apple Silicon

    We wrote our inference engine on Rust, it is faster than llama cpp in all of the use cases. Your feedback is very welcomed. Written from scratch with idea that you can add support …

  4. comment
  5. story
  6. comment
  7. comment
  8. story
  9. comment
    Comment #43553251

    I made it! 90 t/s on my iPhone with llama1b fp16 We completely rewrite the inference engine and did some tricks. This is a summarization with llama 3.2 1b float16. So most of the t…

  10. story
  11. story
  12. story
  13. story
  14. story
  15. story
  16. comment
    Comment #8669306

    It's my first experience on Medium. I hope community can help me to improve my skills.

  17. story
  18. comment
    Comment #8236777

    really great idea, but why morning? %)