I've had to give this some thought for other reasons, and after a couple decades solving analogous problems to moderation in security, I agree with yishan about signal to noise over the specific content, but what I have effectively spent a career studying and detecting with data is a single factor: malice. It's something every person is capable of, and it takes a lot of exercise and practice with higher values to rea…
That is a big stretch. Hate can't be applied to many things, including disagreements like this comment.
But it can be pretty clearly applied to statements that, if carried out in life, would deny another person or peoples' human rights. Another is denigration or mocking someone on the basis of things that can't or shouldn't have to change about themselves, like their race or religion. There is a pretty bright line there.
Malice (per the conventional meaning of something bad intended, but not necessarily revealed or acted out) is a much lower bar that includes outright hate speech.
> but really, you can see when someone is actuated by it.
How can you identify this systematically (vs it being just your opinion), but not identify hate speech?