Live data from Hacker News

Viewing profile — naasking

naasking

HN member
Joined
Fri, Sep 09, 2011, 2:29 PM UTC
HN karma
12,667
Public activity
8,642 items

About naasking

No profile information was provided.

Recent public activity

  1. comment
    Comment #49245331

    People say Qwen overthinks because they analyzed the thinking traces, and Qwen finds the answer relatively quickly but then second guesses itself multiple times for another 20,000+…

  2. comment
    Comment #49212723

    The continued vagueness suggests this isn't going anywhere, so I'll just conclude by saying that you've given nobody any real reason to believe this.

  3. comment
    Comment #49198528

    No such thing has been shown.

  4. comment
    Comment #49198411

    Typical LLM pretraining is very inefficient with data. NanoGPT slowrun shows that data can be used much more efficiently.

  5. comment
    Comment #49196951

    I don't understand all this dismissal of Einstein. What you just described is that he published a unified theory of Brownian motion, developed a new way of conceptualizing scientif…

  6. comment
    Comment #49196170

    I use LLMs regularly for real work. I understand the current limitations, but you're talking about existing limitations that will never be overcome, because they cannot be overcome…

  7. comment
    Comment #49193942

    Yes, it's worse because processors clocks are much faster than memory clocks.

  8. comment
    Comment #49188099

    Yes. That's why law is a different discipline from linguistics requiring its own rules of analysis, pedagogy, etc.

  9. comment
    Comment #49188088

    That's just vibes then. How is that supposed to be convincing?

  10. comment
    Comment #49185768

    > but there are many things which they are not good at which it is not cost effective or meaningful to improve Can you name a few such things so I can keep an eye on them in the co…

  11. comment
    Comment #49185732

    > We know exactly how attention layers work and how they produce the next word as well as draw them from larger feature spaces. This is not what's meant by the statements that we d…

  12. comment
    Comment #49185690

    The search space is far too large for a mere order of magnitude to make any difference at all.

  13. comment
    Comment #49185654

    > we know everything about how LLMs work No we don't. That we understand the low level mechanics of a system doesn't mean we understand how any high level phenomena emerge from tho…

  14. comment
    Comment #49185587

    I don't think that's correct. Increasing parameter count increases capabilities in all domains per the scaling laws. Models are larger than they were years ago, so capabilities in …

  15. comment
    Comment #49180316

    Bribes are not speech though, and advertising is.

  16. comment
    Comment #49176445

    It's interesting to claim that it would be difficult to explain why we punish successful crime more. A successful murder creates more suffering (victim's friends and family), and o…

  17. comment
    Comment #49167994

    It does matter though. If you want to murder someone by hitting them with a plushie, you're not going to get charged with attempted murder because it's not possible that that would…

  18. comment
    Comment #49166693

    The margins for groceries are objectively thin. The only way to provide food at lower prices is to provide a worse good or service, eg. less variety, less quality, less availabilit…

  19. comment
    Comment #49166530

    Yes, but there is little evidence this had a meaningful effect on votes.

  20. comment
    Comment #49166195

    They're important everywhere of course, but especially on mobile. If AI researchers figure out how to offload knowledge and expertise from reasoning weights, then a core reasoning …

  21. comment
    Comment #49142479

    > I can't see how the model could include the actual subjective human experience. People who say LLMs have subjective experience aren't saying they have human-type subjective exper…

  22. comment
    Comment #49142437

    Even mechanistic models generate interesting discussion. How many years have we discussed Turing machines and the lambda calculus? Almost a century of great work came out of those.…

  23. comment
    Comment #49136060

    > Dunno about the parent commenter, but I personally interpret the concept as having a hidden representation of self that is continually tended to I don't see why an LLM could not …

  24. comment
    Comment #49135984

    Making definitive claims about whether LLMs do or do not have specific properties absolutely does require precise definitions of those properties that can be used to evaluate those…

  25. comment
    Comment #49135944

    > The qualia themselves, even those that are quite abstract, are rooted in our physical presence and evolution. There is no objective evidence of qualia. All evidence of qualia are…