Earlier quoted context omitted.
> If I want to ask a software tool what the suicide rate is for my county, I do not expect it to come back with: "Naughty boy! You said an unsafe word! You're getting a strike, and if you get two more, you're banned." Did this happen? I just tested this query in Grok, Gemini, Claude, and ChatGPT and 0% of them admonished me or refused to return an answer. Just like every single conversation I've ever had on this topi…
That's why I said: > Replace "suicide" with whatever the "AI Safety" obsession word is today I don't know what those queries are, but original-OP made one and got a "strike", which is what spawned this thread.
OP gave an example of reverse engineering, something that to the LLM looks identical to just hacking. I am totally fine if the incredibly tiny little fraction of people who want to reverse engineer their own systems can't use LLMs to do it, and in exchange top LLMs aren't helpful for the hordes of actual malicious actors who would love a superintelligence to aid their crimes.
No-brainer tradeoff, just like 100% of examples I've ever heard.