Live data from Hacker News

Viewing profile — mbowcut2

mbowcut2

HN member
Joined
Wed, Feb 09, 2022, 11:53 PM UTC
HN karma
217
Public activity
72 items

About mbowcut2

No profile information was provided.

Recent public activity

  1. comment
    Comment #48692427

    I made myself reason it out, and came up with the exact same intuition. You need a sequence of 2n moves (n down moves, n over moves), but the sequence is completely determined by w…

  2. comment
    Comment #48465780

    Totally agree. Granting exemptions feels like trying to have their cake and eat it too. If regulations mean anything they need to be enforced so we can see the real downstream effe…

  3. comment
    Comment #47220074

    Gotta hit that docker system prune -a

  4. comment
    Comment #47040087

    Loved him in Secondhand Lions.

  5. comment
    Comment #46800201

    If you thought we were getting bad bugs before, just wait until the 90% agent-coded PRs start landing. We're gonna have multiple crowdstrike-level blowups.

  6. comment
    Comment #46737963

    It's an interesting concept, but I'm skeptical about how feasible this is. How much design/legwork/intervention will Seth actually contribute during the entire process? I'm thinkin…

  7. comment
    Comment #46627258

    Wow, I didn't know about the "skills" feature, but with that as context isn't this attack strategy obvious? Running an unverified skill in Cowork is akin to running unverified code…

  8. comment
    Comment #46124029

    It makes me wonder about the gaps in evaluating LLMs by benchmarks. There almost certainly is overfitting happening which could degrade other use cases. "In practice" evaluation is…

  9. comment
    Comment #45531443

    Seems like the less sexy headline is just something about the sample size needed for LLM fact encoding That's honestly a more interesting angle to me: How many instances of data X …

  10. comment
    Comment #45215205

    I'm not surprised. People really thought the models just kept getting better and better?

  11. story
  12. story
  13. comment
    Comment #44827917

    it looks like the 2nd and 3rd bar never got updated from the dummy data placeholders lol.

  14. comment
    Comment #44800161

    It's not a new problem (for individuals), though perhaps at an unprecedented scale (so, maybe a new problem for civilization). I'm sure there were black smiths that felt they had l…

  15. comment
    Comment #44727010

    I've had similar experiences with vanilla ChatGPT as a DM but I bet with clever prompt engineering and context window management you could solve or at least dramatically improve th…

  16. comment
    Comment #44713256

    You can, and there has been some interesting work done with it. The technique is called LogitLens, and basically you pass intermediate embeddings through the LMHead to get logits c…

  17. comment
    Comment #44711582

    The problem with embeddings is that they're basically inscrutable to anything but the model itself. It's true that they must encode the semantic meaning of the input sequence, but …

  18. comment
    Comment #44683244

    LLMs are better at LaTeX than humans. ChatGPT often writes LaTeX responses.

  19. comment
    Comment #44505203

    I think I agree with you. My only rebuttal would be it's this kind of thinking that's kept any leading players form trying other architectures in the first place. As far as I know,…

  20. comment
    Comment #44302881

    I read this as "pirate space industry" and got real excited.

  21. comment
    Comment #44302853

    It's interesting how I couldn't tell whether the rocket was 1m tall or 10m tall in this video. Turns out it's actually 6m tall per the link.

  22. comment
    Comment #44227159

    Nah, I think they made it model agnostic, which is kinda smart.

  23. comment
    Comment #44042969

    To a topologist, everything is topology.

  24. comment
    Comment #43908541

    Pack it up boys, they finally made the killer app.

  25. comment
    Comment #43364838

    I had a smart TV that gradually got slower and slower until it became basically useless. I figured it was just running out of RAM as apps got larger with updates over the years.