Viewing profile — remilouf
remilouf
HN member- Joined
- Sun, May 14, 2017, 7:50 AM UTC
- HN karma
- 501
- Public activity
- 107 items
- HN profile
- View on Hacker News ↗
About remilouf
Recent public activity
-
comment
Comment #48298798
Rémi here, really cool! It kind of turned into a rabbit hole on my end: https://github.com/rlouf/sigil
-
comment
Comment #48279174
Of course: https://github.com/rlouf/sigil
-
comment
Comment #48279167
It was indeed inspired from my IRC days :)
-
comment
Comment #48279156
Original author here, the project has evolved quite a bit since then, you can follow here if that interests you: https://github.com/rlouf/sigil (The $7k was sarcasm)
- story
-
comment
Comment #48047218
Author here. Sorry my writing is tedious. Next time I’ll use AI to make it more readable.
- story
- story
- story
-
comment
Comment #47776539
> Ironically LLMs solve the MxN problem he's complaining about Enlighten me please
-
comment
Comment #47768297
Ooops sorry
-
comment
Comment #47767575
Author here. You're right, it's not a hard problem, but a particularly annoying one.
- story
-
comment
Comment #47626608
I haven't always done this, and the knowledge base used to visibly degrade over time. Reviewing a PR does not take a long time, maybe a few minutes, and this compounds over time.
- story
- story
-
comment
Comment #41895854
This is actually pretty funny.
-
comment
Comment #41504705
That’d be a pretty inefficient way to generate bullshit at scale
- story
-
comment
Comment #40237005
LLM evaluations are very sensitive to the details of the prompt's structure. This post shows how using structured generation reduces the results' variance and the ranking shifts.
- story
- story
-
comment
Comment #39944856
Looks like it’s quite the opposite: http://blog.dottxt.co/performance-gsm8k.html
-
comment
Comment #39716071
What do you mean by "semantic dimension"?
-
comment
Comment #39715873
That whole structured generation line of work looks promising. I hope someone else takes this and runs evaluations on other benchmarks. Curious to see if the results translate!