Live data from Hacker News

Viewing profile — hislaziness

hislaziness

HN member
Joined
Mon, Jun 29, 2009, 7:40 AM UTC
HN karma
478
Public activity
121 items

About hislaziness

No profile information was provided.

Recent public activity

  1. comment
    Comment #48105407

    Am I missing something here. Not finding a major bug/vulnerability just means that maybe the code is really good, not that the model is not what is claimed?

  2. story
  3. story
  4. comment
    Comment #44871577

    Would it be more appropriate to compare LLMs to Autotunes rather than pianos?

  5. comment
    Comment #44641750

    Akamai CTO Robert Blumofe offers four tips for business leaders striving to foster AI fluency by empowering employees with the right tools and best use cases. useful insights but d…

  6. comment
    Comment #44621230

    Great. I enjoy these kind of articles. My all time favorite book for 'C' is Expert C Programming: Deep C Secrets.

  7. comment
    Comment #44621180

    Terence Tao on the matter - https://imgur.com/a/terence-tao-on-supposed-gold-imo-sMKP0bm

  8. comment
    Comment #41921983

    As I understand, the LLM uses the techniques of searchformer - https://arxiv.org/abs/2402.14083 . To do "slow thinking" doing a A* search using a transofrmer.

  9. story
  10. comment
    Comment #41805991

    It is not just MSRP, management and operations cost too. The article goes into the details of this.

  11. comment
    Comment #41805949

    The details are in the article. They have done the math.

  12. comment
    Comment #41805550

    TLDR: Don’t buy H100s. The market has flipped from shortage ($8/hr) to oversupplied ($2/hr), because of reserved compute resales, open model finetuning, and decline in new foundati…

  13. comment
    Comment #41764325

    The screen seems to be stuck at Wait a second... for me.

  14. comment
    Comment #41470901

    I tried a few local LLMs. None of them could give me the right answer for "How many 'r's in straberry. All LLMs were 8-27B.

  15. comment
    Comment #41343607

    I know you mean this in jest, but we are much closer to this than we would imagine, the use of LLMs to process communication / translation is becoming ubiquitous. We are 1 bad tran…

  16. comment
    Comment #41253755

    I also use some email providers ability to have +xyz at the end of the username. So for a registration I would for user.name+sitea@domain.com. has helped me track spam and leaks in…

  17. comment
    Comment #41053839

    Cool. I will try it out. I tried the same with ollama, the non english part needs a lot more polish. Do you see the outcome being any different?

  18. comment
    Comment #41053081

    This is pretty awesome. How did you do it? Any blog on the detials?

  19. comment
    Comment #41000896

    The model description on huggingface says - Model size - 12.2B params, Tensor type - BF16. Is the Tensor type different from the training param size?

  20. comment
    Comment #41000885

    I just checked huggingface and the model files download is about 25GB but in a comment below someone mentioned it is 8fp quantized model. Trying to understand how the quantization …

  21. comment
    Comment #40997003

    isn't it 2 bytes (fp16) per param. so 7b = 14 GB+some for inference?

  22. comment
    Comment #40968186

    The title does not do justice to the article. It talks about a new classification system that OpenAI has introduced for LLMs

  23. story
  24. comment
    Comment #40799403

    [flagged]

  25. comment
    Comment #40605266

    Same here. Coincidentally I joined in June 2009 and this is the only social media I indulge in.