Live data from Hacker News

Viewing profile — aarnphm

aarnphm

HN member
Joined
Tue, May 17, 2022, 8:48 PM UTC
HN karma
78
Public activity
14 items

About aarnphm

ml and system

Recent public activity

  1. story
  2. story
  3. comment
    Comment #44550121

    Hi srameshc, the core of BentoML is still considered MLOps. A lot of our customers are pretty much MLOps users. However, LLMOps seem like a natural progression of the product, give…

  4. comment
    Comment #44550110

    Thanks for the recommendation, I'm actually working on something similar for this part of the docs (I'm also working at BentoML).

  5. story
  6. comment
    Comment #37233164

    You can pretty much make the same argument about Docker. Docker abstracts away runc, runc abstract away cgroup. I don't think calling it abstraction is correct. OneDiffusion is des…

  7. story
  8. story
  9. comment
    Comment #36393230

    Currently on main, 8bit and 4bit quant is supported One can simply do ```openllm start falcon --model-id tiiuae/falcon-40b-instruct --quantize int4``` Beware that there is no free …

  10. comment
    Comment #36388895

    Hi there, 8bit and 4bit is currently supported on main. GPTQ is working in progress, as well as GGML

  11. comment
    Comment #36388774

    Hi all, I'm the main maintainer from the OpenLLM team here. I'm actively developing the fine-tuning feature and will release a PR soon enough. Stay tuned. In the meanwhile, the bes…

  12. comment
    Comment #34444161

    Again?? I definitely have to try out BentoML then to see what the hype is about!!

  13. comment
    Comment #32088676

    BentoML is an amazing tool that allows you to quickly and easily deploy your machine learning models as APIs. It is extremely user-friendly and easy to use, and the results are ama…

  14. comment