Live data from Hacker News

Viewing profile — remilouf

remilouf

HN member
Joined
Sun, May 14, 2017, 7:50 AM UTC
HN karma
501
Public activity
107 items

About remilouf

Twitter: @remilouf GH: https://github.com/rlouf

Recent public activity

  1. comment
    Comment #48298798

    Rémi here, really cool! It kind of turned into a rabbit hole on my end: https://github.com/rlouf/sigil

  2. comment
    Comment #48279174

    Of course: https://github.com/rlouf/sigil

  3. comment
    Comment #48279167

    It was indeed inspired from my IRC days :)

  4. comment
    Comment #48279156

    Original author here, the project has evolved quite a bit since then, you can follow here if that interests you: https://github.com/rlouf/sigil (The $7k was sarcasm)

  5. story
  6. comment
    Comment #48047218

    Author here. Sorry my writing is tedious. Next time I’ll use AI to make it more readable.

  7. story
  8. story
  9. story
  10. comment
    Comment #47776539

    > Ironically LLMs solve the MxN problem he's complaining about Enlighten me please

  11. comment
    Comment #47768297

    Ooops sorry

  12. comment
    Comment #47767575

    Author here. You're right, it's not a hard problem, but a particularly annoying one.

  13. story
  14. comment
    Comment #47626608

    I haven't always done this, and the knowledge base used to visibly degrade over time. Reviewing a PR does not take a long time, maybe a few minutes, and this compounds over time.

  15. story
  16. story
  17. comment
    Comment #41895854

    This is actually pretty funny.

  18. comment
    Comment #41504705

    That’d be a pretty inefficient way to generate bullshit at scale

  19. story
  20. comment
    Comment #40237005

    LLM evaluations are very sensitive to the details of the prompt's structure. This post shows how using structured generation reduces the results' variance and the ranking shifts.

  21. story
  22. story
  23. comment
    Comment #39944856

    Looks like it’s quite the opposite: http://blog.dottxt.co/performance-gsm8k.html

  24. comment
    Comment #39716071

    What do you mean by "semantic dimension"?

  25. comment
    Comment #39715873

    That whole structured generation line of work looks promising. I hope someone else takes this and runs evaluations on other benchmarks. Curious to see if the results translate!