"Crucially, it tells the agent not to rely on its internal training data (which might be hallucinated or refer to a different version of the game) but to ground its knowledge in what it observes. " Does this even have any effect?
Whether the 'effect' something implied by the prompt, or even something we can understand, is a totally different question.