Live data from Hacker News

Viewing profile — p1esk

p1esk

HN member
Joined
Sat, May 26, 2012, 11:19 PM UTC
HN karma
6,511
Public activity
4,433 items

About p1esk

hncomments5@gmail.com

Recent public activity

  1. comment
    Comment #49204141

    If Siri is using a 3T model in high reasoning mode to answer your question you will.

  2. comment
    Comment #49163957

    I stopped writing code completely about a year ago, stopped reading code completely about 8 months ago, and I feel like I stopped thinking hard about anything at work about 6 month…

  3. comment
    Comment #49161167

    Is there any degradation with INT8 weights quantization? Why would anyone want to apply ConvRot to do 8 bit weights? Note the paper [1] focuses on 4 bit weights and 4 bit activatio…

  4. comment
    Comment #49160537

    Zero. These were open problems.

  5. comment
    Comment #49114989

    This blog post

  6. comment
    Comment #49111585

    It’s still at GPT-1 level, but GPT-2 moment feels imminent.

  7. comment
    Comment #49087470

    replace they key-query-value mechanic by just dropping it while making the entire context the latent space. What do you mean by this? Like concatenating all token embeddings into o…

  8. comment
    Comment #49051203

    I wonder how difficult it would be to convert these to 3D

  9. comment
    Comment #49047278

    I’m guessing other models would probably stop in this situation and ask the user for instructions.

  10. comment
    Comment #49046161

    I think they meant that Opus 5 had to find and set up a vision model to process the image

  11. comment
    Comment #48975025

    [flagged]

  12. comment
    Comment #48969601

    [flagged]

  13. comment
    Comment #48877272

    It’s nice to be rich I guess

  14. comment
    Comment #48808232

    where do you see "twice the memory bandwidth"?

  15. comment
    Comment #48781717

    There’s noticeable accuracy degradation when they switched from fp8 to mxfp4

  16. comment
    Comment #48751979

    That’s how you get skills

  17. comment
    Comment #48748816

    Do a decentralized p2p one

  18. comment
    Comment #48728098

    We’re experiencing gpt-2 moment in robotics now. This means in about 2-3 years they will do useful work (cooking, repairs, cleaning, etc).

  19. comment
    Comment #48703018

    It’s got to be similar to Fable, which I experienced for 3 days, and which impressed me (compared to Opus 3.8)

  20. comment
    Comment #48702967

    VLAs are new LLMs. Give them 5 years to develop. But even good old LLMs are still improving every six months.

  21. comment
    Comment #48702735

    Unfortunately if your manager thinks something is your problem, it becomes your problem.

  22. comment
    Comment #48701717

    You realize llms as a field is barely 5 years old? Give it at least another 5.

  23. comment
    Comment #48625734

    Many people here spent a lot more than $300 on headphones long before AirPods appeared.

  24. comment
    Comment #48566726

    100×10^15 km

  25. comment
    Comment #48520374

    Do you feel that recent advances in AI can speed up such rare disease research?