Live data from Hacker News

Viewing profile — ipiszy

ipiszy

HN member
Joined
Mon, Oct 17, 2011, 10:04 AM UTC
HN karma
11
Public activity
10 items

About ipiszy

developer

Recent public activity

  1. comment
    Comment #33091571

    For now it's for single GPU inference only.

  2. comment
    Comment #33091549

    RTX 3080-10GB should work. You could check https://github.com/facebookincubator/AITemplate/tree/main/ex... , and https://www.reddit.com/r/StableDiffusion/comments/xv7m89/met... .

  3. comment
    Comment #33085256

    Yes this is correct. batch 16 7.9s / 25 steps, per image 0.49s: it generates 16 images for each prompt within 7.9s, so it's 0.49s per image.

  4. comment
    Comment #33085204

    AITemplate only supports fp16 data types with fp16 or fp32 accumulation right now. We are working on supporting more data types and quantization. We don't have an official comparis…

  5. comment
    Comment #33084609

    We have a bunch of unittests and E2E tests to compare numeric numbers between AITemplate and PyTorch eager.

  6. comment
    Comment #33084598

    You could check "AITemplate optimizations" section in the blog ( https://ai.facebook.com/blog/gpu-inference-engine-nvidia-amd... ), and https://github.com/facebookincubator/AITempl…

  7. comment
    Comment #33076799

    As @haolu7 mentioned, you could take a pre-trained model and use AITemplate to do model inference. All you need to do is to re-write the model using AITemplate frontend and map PyT…

  8. comment
    Comment #33073176

    tl;dr: Meta is open sourcing AITemplate, an inference engine for both Nvidia and AMD GPUs. Code: https://github.com/facebookincubator/AITemplate . AITemplate delivers much better p…

  9. comment
    Comment #5077266

    I like HackerNode more because: 1. It contains more columns than only "FrontPage", such as "Jobs" and "Comments"; 2. The UI seems more concise. Although there are not indent betwee…

  10. comment
    Comment #3783070

    ...............................