Live data from Hacker News

Viewing profile — accountnum

accountnum

HN member
Joined
Thu, Sep 05, 2024, 8:38 AM UTC
HN karma
12
Public activity
20 items

About accountnum

No profile information was provided.

Recent public activity

  1. comment
    Comment #41539650

    [flagged]

  2. comment
    Comment #41539568

    [flagged]

  3. comment
    Comment #41539019

    [flagged]

  4. comment
    Comment #41537541

    It's not a problem, because the point at which we are in the logarithmic curve is the only thing that matters. No one in their right mind ever expected anything linear, because tha…

  5. comment
    Comment #41536026

    I'm going to simply address what I think are your main points here. There is nowhere that an LLM stores all possible outputs. Causality can trivially be represented by sampling by …

  6. comment
    Comment #41535380

    Again, you're stretching definitions into meaninglessness. The way you are using "sampling" and "distribution" here applies to any system processing any information. Yes, humans as…

  7. comment
    Comment #41535268

    They're not sampling from prior conversations. The model constructs abstracted representations of the domain-specific reasoning traces. Then it applies these reasoning traces in va…

  8. comment
    Comment #41533206

    My point isn't that the model falls for gender stereotypes, but that it falls for thinking that it needs to solve the unmodified riddle. Humans fail at the original because they ex…

  9. comment
    Comment #41531195

    Recognizing that it is a riddle isn't impressive, true. But the duration of its reasoning is irrelevant, since the riddle works on misdirection. As I keep saying here, give someone…

  10. comment
    Comment #41530211

    The trick with the 7 wives and 7 bags and so on is that no long reasoning is required. You just have to notice one part of the question that invalidates the rest and not shortcut t…

  11. comment
    Comment #41530172

    No, that's not a conclusion we can draw, because there is nothing much more to do than memorize the answer to this specific trick question. That's why it's a trick question, it goe…

  12. comment
    Comment #41529787

    It hasn't read that riddle because it is a modified version. The model would in fact solve this trivially if it _didn't_ see the original in its training. That's the entire trick.

  13. comment
    Comment #41529772

    It literally is a riddle, just as the original one was, because it tries to use your expectations of the world against you. The entire point of the original, which a lot of people …

  14. comment
    Comment #41529731

    No, it's necessary to either know that it's a trick question or to have a feeling that it is based on context. The entire point of a question like that is to trick your understandi…

  15. comment
    Comment #41505311

    [flagged]

  16. comment
    Comment #41503765

    [flagged]

  17. comment
    Comment #41503658

    No, you're not. Are you genuinely trying to suggest that LLMs, which can: - Construct arbitrary text that isn't just grammatically but semantically coherent - Derive intent, subtle…

  18. comment
    Comment #41460335

    A weakness of the current models in some domains considered useful, yes - but not a fundamental limitation of the architecture. I see no consensus on the latter whatsoever. The ARC…

  19. comment
    Comment #41458863

    > One reason why just blathering on endlessly... First of all, I would urge you to stop arbitrarily using negative words to make an argument. Saying that LLMs are "blathering" is e…

  20. comment
    Comment #41454894

    You seem to repeatedly insist that hidden computation is a distinction of any relevance whatsoever. First of all, your understanding of the architecture itself is mistaken. A trans…