Any ideas how to solve the agent's don't have total common sense problem? I have found when using agents to verify agents, that the agent might observe something that a human would immediately find off-putting and obviously wrong but does not raise any flags for the smart-but-dumb agent.
Architecturally focusing on Episodic memory with feedback system.
This training is retrieved next time when something similar happens