Live data from Hacker News

Watching o3 model sweat over a Paul Morphy mate-in-2

alexop.dev

1–10 of 73 posts

Re: Watching o3 model sweat over a Paul Morphy mate-in-2

#8

On a similar note, I just updated LLM Chess Puzzles repo [1] yesterday. The fact that gpt-4.5 gets 85% correctly solved is unexpected and somewhat scary (if model was not trained on this). [1] https://github.com/kagisearch/llm-chess-puzzles

Oh cool, I wonder how good 03 will be. While using 03, I noticed something funny: sometimes I gave it a screenshot without any position data. It ended up using Python and spent 10 minutes just trying to figure out where the figures were exactly.
Post reply on HN