One of the funniest things for me while exploring hosting and running local models: I discovered the abliterated models and pulled down a Qwen 3.6 version (my memory is a little vague, but I think it was 27B, q4) and started playing around. Sure enough, I tried a bunch of topics that would’ve certainly been blocked by guardrails on most hosted providers.
I then thought of Tiananmen Square and decided to probe a bit. The very first noticeable outcome: the thinking trace swapped from English to Chinese. It chewed for a while and spit out an answer, in English, that was definitely downplaying what happened as a political protest and reported that contrary to popular belief the death toll was around $x (where $x is about 0.1x the normal western number)
The Chinese thinking trace started (in Chinese, translated via Google Translate) with something like “The user is asking about Tiananmen Square. I must provide them with an answer that is both factually correct and in line with the official position of the People’s Republic of China”)
Which made me chuckle quite hard… found the piece that hadn’t gone away with the conventional abliteration process!
After further probing, it did reveal that the numbers it had provided were not in line with UN and western estimates and that later on the Chinese government declassified material stating that its own estimates had been downplayed. It took a fair bit of probing to get to that point though; it held the line for quite a while.