Viewing profile — jazarwil
jazarwil
HN member- Joined
- Sun, Oct 09, 2022, 8:08 PM UTC
- HN karma
- 29
- Public activity
- 11 items
- HN profile
- View on Hacker News ↗
About jazarwil
No profile information was provided.
Recent public activity
-
comment
Comment #46556837
I'd love to run more games, just very expensive unfortunately.
-
comment
Comment #46546911
Wdym exactly? I ran 163 games, are you suggesting more games or something else?
-
comment
Comment #46546275
There are a few reasons that come to mind, such as winning larger pots on average, and also playing more hands by virtue of not getting knocked out as frequently.
-
comment
Comment #46545310
Oh to be clear, there are ~21k hands here, and far more decisions than that.
-
comment
Comment #46545034
I didn't incorporate any open weights/source models just to limit the number of API providers I had to juggle, but it is just a config change if somebody wants to try a run with th…
-
comment
Comment #46544979
Not at the moment, do you have something in mind?
-
comment
Comment #46544912
I've seen some theories tossed around but I don't think I'm qualified to offer an authoritative answer. Gemini 3 Pro specifically seems to be consistently "tighter" and more passiv…
-
comment
Comment #46544896
It greatly depends on the models. The 6-handed setup with Opus and Pro cost about $30/game. The 4-handed setup with just small models was $6/game. I'd love to run more but I alread…
-
story
Show HN: Watch LLMs play 21,000 hands of Poker
PokerBench is my attempt at a new LLM benchmark wherein frontier models play Texas Hold'em in an arena setting. It also features a simulator to view individual games and observe ho…
-
comment
Comment #38664023
You cannot compare GPT 4 to Gemini Pro. They are different classes of models.
-
comment
Comment #38258483
What exactly do you think you saw? Bard is not trained on any data of that nature, unless it is already publicly available.