Live data from Hacker News

Viewing profile — antinucleon

antinucleon

HN member
Joined
Wed, Apr 09, 2014, 3:33 AM UTC
HN karma
118
Public activity
34 items

About antinucleon

No profile information was provided.

Recent public activity

  1. story
  2. story
  3. story
  4. story
  5. story
  6. story
  7. story
  8. story
  9. story
  10. comment
  11. comment
    Comment #36169483

    Mojo is trying to create a new language to solve the problem, and specialized for CPU. We are using a more pragmatic way to solve GPU AI computation problem.

  12. comment
    Comment #36169437

    We haven't compared yet.

  13. comment
    Comment #36169429

    We developed AITemplate majorly for Meta's focus at that time, eg Ads/Ranking need. For HippoML is startup we are building for Generative AI. HippoML is not using AITemplate.

  14. comment
    Comment #36169315

    It is actually non-trivial to get GPU run fast, especially on SoC with strong CPU like M2.

  15. comment
    Comment #36169032

    Hippo is faster than AITemplate, and supports more generative models. We haven't compared vs TVM, but for absolute token/s on M2 Max, Hippo is able to run decoding on LLAMA with da…

  16. comment
    Comment #36168947

    We will disclose more details very soon.

  17. comment
    Comment #36168936

    Yes. We support >= 1bit <= 16bit models out of box for various of models.

  18. comment
    Comment #36168666

    AITemplate's original designer is here. We quit Meta in January and start HippoML ( https://hippoml.com/ ). We just disclosed our new engine's performance on LLM: https://blog.hipp…

  19. comment
    Comment #17129690

    CuDNN v7 was used in the experiments, in the experiments parts each comparison was listed with version or commit number.

  20. comment
    Comment #17128056

    Summary: Tensor program is able to be optimized by using machine learning and transfer learning. The numerical program optimization model is trained on feature from low-level AST o…

  21. story
  22. comment
    Comment #13020349

    The version tested in paper should not have P2P,so it was much slower than current version.

  23. comment
    Comment #13018491

    There is a huge distributed performance advantages vs TensorFlow. You can get a hint from Prof. Carlos Guestrin's keynote talk at Data Science Summit 2016. Also, CMU CS Dean Andrew…

  24. story
  25. story