It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.
Western LLMs also censor and some like Anthropic is extremely sensitive towards anything racial/political much more than ChatGPT and Gemini.
The golden chalice is an uncensored LLM that can run locally but we simply do not have enough VRAM or a way to decentralize the data/inference that will remove the operator from legal liability.