Viewing profile — aarnphm
aarnphm
HN member- Joined
- Tue, May 17, 2022, 8:48 PM UTC
- HN karma
- 78
- Public activity
- 14 items
- HN profile
- View on Hacker News ↗
About aarnphm
Recent public activity
- story
- story
-
comment
Comment #44550121
Hi srameshc, the core of BentoML is still considered MLOps. A lot of our customers are pretty much MLOps users. However, LLMOps seem like a natural progression of the product, give…
-
comment
Comment #44550110
Thanks for the recommendation, I'm actually working on something similar for this part of the docs (I'm also working at BentoML).
- story
-
comment
Comment #37233164
You can pretty much make the same argument about Docker. Docker abstracts away runc, runc abstract away cgroup. I don't think calling it abstraction is correct. OneDiffusion is des…
- story
- story
-
comment
Comment #36393230
Currently on main, 8bit and 4bit quant is supported One can simply do ```openllm start falcon --model-id tiiuae/falcon-40b-instruct --quantize int4``` Beware that there is no free …
-
comment
Comment #36388895
Hi there, 8bit and 4bit is currently supported on main. GPTQ is working in progress, as well as GGML
-
comment
Comment #36388774
Hi all, I'm the main maintainer from the OpenLLM team here. I'm actively developing the fine-tuning feature and will release a PR soon enough. Stay tuned. In the meanwhile, the bes…
-
comment
Comment #34444161
Again?? I definitely have to try out BentoML then to see what the hype is about!!
-
comment
Comment #32088676
BentoML is an amazing tool that allows you to quickly and easily deploy your machine learning models as APIs. It is extremely user-friendly and easy to use, and the results are ama…
- comment