Viewing profile — ipiszy
ipiszy
HN member- Joined
- Mon, Oct 17, 2011, 10:04 AM UTC
- HN karma
- 11
- Public activity
- 10 items
- HN profile
- View on Hacker News ↗
About ipiszy
Recent public activity
-
comment
Comment #33091571
For now it's for single GPU inference only.
-
comment
Comment #33091549
RTX 3080-10GB should work. You could check https://github.com/facebookincubator/AITemplate/tree/main/ex... , and https://www.reddit.com/r/StableDiffusion/comments/xv7m89/met... .
-
comment
Comment #33085256
Yes this is correct. batch 16 7.9s / 25 steps, per image 0.49s: it generates 16 images for each prompt within 7.9s, so it's 0.49s per image.
-
comment
Comment #33085204
AITemplate only supports fp16 data types with fp16 or fp32 accumulation right now. We are working on supporting more data types and quantization. We don't have an official comparis…
-
comment
Comment #33084609
We have a bunch of unittests and E2E tests to compare numeric numbers between AITemplate and PyTorch eager.
-
comment
Comment #33084598
You could check "AITemplate optimizations" section in the blog ( https://ai.facebook.com/blog/gpu-inference-engine-nvidia-amd... ), and https://github.com/facebookincubator/AITempl…
-
comment
Comment #33076799
As @haolu7 mentioned, you could take a pre-trained model and use AITemplate to do model inference. All you need to do is to re-write the model using AITemplate frontend and map PyT…
-
comment
Comment #33073176
tl;dr: Meta is open sourcing AITemplate, an inference engine for both Nvidia and AMD GPUs. Code: https://github.com/facebookincubator/AITemplate . AITemplate delivers much better p…
-
comment
Comment #5077266
I like HackerNode more because: 1. It contains more columns than only "FrontPage", such as "Jobs" and "Comments"; 2. The UI seems more concise. Although there are not indent betwee…
-
comment
Comment #3783070
...............................