Earlier quoted context omitted.
Pretty sure it gets the entire conversation as input
Is this described somewhere? Wikipedia doesn't help
http://jalammar.github.io/illustrated-gpt2/
It uses both its own output from previous steps and the users prompt(s) as input for each token(word) that it predicts.