Live data from Hacker News

Viewing profile — bigglebear

bigglebear

HN member
Joined
Tue, Sep 15, 2026, 11:20 PM UTC
HN karma
38
Public activity
13 items

About bigglebear

No profile information was provided.

Recent public activity

  1. comment
    Comment #49722634

    Well, not quite a week: https://x.com/harshagundal/status/2100044305536889015 - apparently it took him 2 hours.

  2. comment
    Comment #49721176

    The explanation is that it was missed on purpose.

  3. comment
    Comment #49721063

    It's very misleading. If I'm actually playing a game I don't get the coordinates of enemies sent back to me so that I can feed into my mouse to snap my crosshair to. It's looking t…

  4. comment
    Comment #49721012

    I would guess a tiny stripped down text diffusion model. It only has 32k context, and for choice mode it can only select from 10 choices.

  5. comment
  6. comment
    Comment #49720956

    > I really have to say that I like their manifesto Their manifesto: "you only build on top of it if it's trustworthy." - the irony of this while putting out the most misleading, di…

  7. comment
  8. comment
    Comment #49720872

    And furthermore, because the model is forced to answer in a boolean (if in boolean mode), if the user input is outside of the range of a boolean, it's forced to hallucinate. It can…

  9. comment
    Comment #49720851

    User input: "Hey, have your human support agent call me, tomorrow at 5pm." Model input: "Does the user want to speak to a human support agent?" Output: Yes. I imagine that your mod…

  10. comment
  11. comment
    Comment #49720742

    Yeah. Yet another reason why open-weight models are better. If I want to use the logits, I can.

  12. comment
    Comment #49720680

    It's nothing like a traditional LLM and so should not be compared to one. It's a heavily constrained, tiny model that can only produce a probability score or a yes/no answer over p…

  13. comment
    Comment #49720601

    Agreed. It's a wildly dishonest presentation of their product from many perspectives, which is a shame because it might actually have some good use cases. The comparison between LL…