Earlier quoted context omitted.
the LLMs would get crushed
To expand on this - an LLM will try to play (and reason) like a person would, while a solver simply crunches the possibility space for the mathematically optimal move. It’s similar to how an LLM can sometimes play chess on a reasonably high (but not world-class) level, while Stockfish (the chess solver) can easily crush even the best human player in the world.
Show HN: Play poker with LLMs, or watch them play against each other
11–20 of 101 posts
Re: Show HN: Play poker with LLMs, or watch them play against each other
#12Earlier quoted context omitted.
the LLMs would get crushed
To expand on this - an LLM will try to play (and reason) like a person would, while a solver simply crunches the possibility space for the mathematically optimal move. It’s similar to how an LLM can sometimes play chess on a reasonably high (but not world-class) level, while Stockfish (the chess solver) can easily crush even the best human player in the world.
Re: Show HN: Play poker with LLMs, or watch them play against each other
#13Earlier quoted context omitted.
the LLMs would get crushed
To expand on this - an LLM will try to play (and reason) like a person would, while a solver simply crunches the possibility space for the mathematically optimal move. It’s similar to how an LLM can sometimes play chess on a reasonably high (but not world-class) level, while Stockfish (the chess solver) can easily crush even the best human player in the world.
Re: Show HN: Play poker with LLMs, or watch them play against each other
#14Re: Show HN: Play poker with LLMs, or watch them play against each other
#15Earlier quoted context omitted.
To expand on this - an LLM will try to play (and reason) like a person would, while a solver simply crunches the possibility space for the mathematically optimal move. It’s similar to how an LLM can sometimes play chess on a reasonably high (but not world-class) level, while Stockfish (the chess solver) can easily crush even the best human player in the world.
How does a poker solver select bet size? Doesn't this depend on posteriors on the opponent's 'policy' + hand estimation?
Re: Show HN: Play poker with LLMs, or watch them play against each other
#16Re: Show HN: Play poker with LLMs, or watch them play against each other
#17Earlier quoted context omitted.
To expand on this - an LLM will try to play (and reason) like a person would, while a solver simply crunches the possibility space for the mathematically optimal move. It’s similar to how an LLM can sometimes play chess on a reasonably high (but not world-class) level, while Stockfish (the chess solver) can easily crush even the best human player in the world.
Unlike Chess, in poker you don’t have perfect information, so there’s no real way to optimize it.
Re: Show HN: Play poker with LLMs, or watch them play against each other
#18I'm not an expert, but as I understand it there are existing solvers for poker/holdem? Perhaps one of the players could be a traditional solver to see how the LLMs fare against those?
Re: Show HN: Play poker with LLMs, or watch them play against each other
#19These LLMs are playing better than most human players I encounter (low limits).
They're kinda bad, but not as criminally bad as the humans.