This is a LLM directly, purposefully lying, i.e. telling a user something it knows not to be true. This seems like a cut-and-dry Trust & Safety violation to me. It seems the LLM is given conflicting instructions: 1. Don't reference memory without explicit instructions 2. (but) such memory is inexplicably included in the context, so it will inevitably inform the generation 3. Also, don't divulge the existence of user-…
The pattern generation engine didn't take into account the prioritized patterns provided by its authors. The tool recognized this pattern in its output and generated patterns that can be interpreted as acknowledgement and correction. Whether this can be considered a failure, let alone a "Trust & Safety violation", is a matter of perspective.