Live data from Hacker News

Prover-Verifier Games improve legibility of language model outputs

openai.com

31–33 of 33 posts

Re: Prover-Verifier Games improve legibility of language model outputs

#31
post #29

Earlier quoted context omitted.

> it frustrates attempts to “detect” if LLMs are used and that is very good. Why is that good?

If you are asking, you’re the kind of person it’s designed to frustrate. Good. Stay frustrated.

Is this an insult, or a criticism?

If you think I’m doing something I shouldn’t, tell me what you think I’m doing that I shouldn’t, and why I shouldn’t?

Why would you just wish me ill?

Re: Prover-Verifier Games improve legibility of language model outputs

#33

Earlier quoted context omitted.

ELI6 why SPAG is better than just the default pretraining method (token context statistics?) of an LLM.

I don't think of SPAG as a replacement for pretraining. For SPAG to work effectively, I would think that it would have to start with an LLM that is pretrained with self-supervised / imitation learning on regular next-token prediction. Think of SPAG as more of a competitor to RLHF than to pretraining. RL is what gave AlphaGo the edge to finally go beyond merely imitating human games, and finally achieve something new.…

nice try, sneaky prover

(thank you)

Post reply on HN