Live data from Hacker News

Viewing profile — snats

snats

HN member
Joined
Thu, Jul 13, 2023, 5:14 AM UTC
HN karma
277
Public activity
91 items

About snats

snats.xyz

Recent public activity

  1. comment
    Comment #48152642

    did a pretty similar thing last month for the text rendering library last month. trained and made a viz for the model and then made it displace text. should probably do a proper wr…

  2. story
  3. comment
    Comment #47434357

    you can also have removed layers of models and keep the same score in benchmarks [1]. i feel that sometimes a lot of the layers might just be redundant and are not fully needed onc…

  4. story
  5. story
  6. comment
    Comment #46623255

    personal website: snats.xyz weblog: weblog.snats.xyz

  7. story
  8. story
  9. story
  10. comment
    Comment #45401341

    I also did a couple of experiments with pruning LLMs[1] using genetic algorithms and you can just keep removing a surprising amount of layers in big models before they start to hav…

  11. story
  12. comment
    Comment #44588535

    i get it. i want one of those. the problem is that most cellphones are not actual cellphones, they are entertainment machines. they are a pocket tv / social media feed place. most …

  13. story
  14. comment
    Comment #43815966

    I am working on the https://moviemovie.club/about , it's a tiny website about film review. It works like a run club, where you have to make a review first to see other people's rev…

  15. story
  16. comment
    Comment #43367415

    Not an insider but imo the work on diffusion language models like LLaDA is really exciting. It's pretty obvious that LLMs are good but they are pretty slow. And in a world where pe…

  17. comment
    Comment #43364489

    It's pretty funny to test in-distribution for AI models. But they fail horribly once you push them a bit[1]. I recently made LLMs play Minesweeper and ALL LLMs that I tested had a …

  18. story
  19. story
  20. story
  21. comment
    Comment #43042652

    yup, if i went to do a PhD interpretability is the only interesting subject for academia IMO right now

  22. story
  23. comment
    Comment #42544882

    It's more of a distilled model, not a fair 1:1 comparison

  24. story
  25. comment
    Comment #42311722

    I use .XYZ because it was pretty cheap when I bought it