In the past humanity people thought of them as this non-emotional purely logic driven beings and in turn problems people imagined would be such as this logic not considering emotions and emotional well being of humans or blindly (but using logic) pursing a specific goal no matter the consequences or not valuing freedom or gaining emotion. But in all that it still follows logic.
But now we have AIs which could have all the problems above _but doesn't use logical thinking_ to _archive goals_. Instead it uses complex overlapping _statistical models_ to _tell a believable story_ where believable is defined by the training data which is _widely inconsistent, wrong, misleading, discriminating, emotionally charged, etc._ because it's just scrapped from the internet. So there is _no systematic finding of goals, subgoals, plan etc_, there is _no logic_, the concept of "truth" simply _doesn't exist_ for such systems etc.
At the same time this turned out to be "often times" good enough to be usable for many task and can be convincing enough to make people believe that it's sentient.
But this also means it will retell common false information, misconceptions, discrimination, hatred etc. from the internet.
Similar it will do what people call "hallucination" and "lying" but it _not_ either of that and calling it that is misleading. Because it just doing _exactly_ what it was created for: Telling a "believable" story given the training data.
And gaslighting, misleading people and lying are a extremely deep ingrained part of the internet, i.e. a deep ingrained part of the data on which it bases what is "believable".
And while we can add tones of bandaid on top to try to hide/filter out such "bad" responses IMHO without fundamentally either changing the training data or the approach this is bound to fail while even stronger upholding a misleading illusion.