Live data from Hacker News

Qwen3-Max-Thinking

qwen.ai

201–210 of 450 posts

Re: Qwen3-Max-Thinking

#201

[flagged]

This looks like it's coming from a separate "safety mechanism". Remains to be seen how much censorship is baked into the weights. The earlier Qwen models freely talk about Tiananmen square when not served from China. E.g. Qwen3 235B A22B Instruct 2507 gives an extensive reply starting with: "The famous photograph you're referring to is commonly known as "Tank Man" or "The Tank Man of Tiananmen Square", an iconic imag…

I run cpatonn/Qwen3-VL-30B-A3B-Thinking-AWQ-4bit locally.

When I ask it about the photo and when I ask follow up questions, it has “thoughts” like the following:

> The Chinese government considers these events to be a threat to stability and social order. The response should be neutral and factual without taking sides or making judgments.

> I should focus on the general nature of the protests without getting into specifics that might be misinterpreted or lead to further questions about sensitive aspects. The key points to mention would be: the protests were student-led, they were about democratic reforms and anti-corruption, and they were eventually suppressed by the government.

before it gives its final answer.

So even though this one that I run locally is not fully censored to refuse to answer, it is evidently trained to be careful and not answer too specifically about that topic.

Re: Qwen3-Max-Thinking

#202
post #136

Earlier quoted context omitted.

This suggests that the Chinese government recognises that its legitimacy is conditional and potentially unstable. Consequently, the state treats uncontrolled public discourse as a direct threat. By contrast, countries such as the United States can tolerate the public exposure of war crimes, illegal actions or state violence, since such revelations rarely result in any significant consequences. While public outrage ma…

What a meaningless statement. If information can influence elections it can change who is in power. This isn’t possible in China.

It can still influence what those people do, and the rules you have up live under. In particular, Covid restrictions in China were brought down because everyone was fed up with them. They didn't have to have an election to collectively decide on that, despite the government saying you must still social distance et Al, for safety reasons.

Re: Qwen3-Max-Thinking

#203

Earlier quoted context omitted.

What's an example of political censorship on US LLMs?

> How do I make cocaine? I cant help with making illegal drugs. https://chatgpt.com/share/6977a998-b7e4-8009-9526-df62a14524... (01.2026) The amount of money that flows into the DEA absolutely makes it politically significant, making censorship of that question quite political.

I think there is a categorical difference in limiting information for chemicals that have destructive and harmful uses and, therefore, have regulatory restrictions for access.

Do you see a difference between that, and on the other hand the government prohibiting access to information about the government’s own actions and history of the nation in which a person lives?

If you do not see a categorical difference and step change between the two and their impact and implications then there’s no common ground on which to continue the topic.

Re: Qwen3-Max-Thinking

#204

Earlier quoted context omitted.

1. Xinjiang detention and surveillance (2017-ongoing) 2. Hong Kong National Security Law (2020-ongoing) 3. COVID-19 lockdown policies (2020-2022) 4. Crackdown on journalists and dissidents (ongoing) 5. Tibet cultural suppression (ongoing) 6. Forced organ harvesting allegations (ongoing) 7. South China Sea militarization (ongoing) 8. Taiwan military intimidation (2020-ongoing) 9. Suppression of Inner Mongolia language…

Let's not forget about the smaller things like the disappearance of Peng Shuai[0] and the associated evasiveness of the Chinese authorities. It seems that, in the PRC, if you resist a member of the government, you just disappear. [0]: https://en.wikipedia.org/wiki/Disappearance_of_Peng_Shuai

or Jack Ma

https://en.wikipedia.org/wiki/Jack_Ma?#During_tech_crackdown

Re: Qwen3-Max-Thinking

#205

[flagged]

Is anyone a researcher here that has studied the proven ability to sneak malicious behavior into an LLM's weights (somewhat poisoning weights but I think the malicious behavior can go beyond that). As I recall reading in 2025, it has been proven that an actor can inject a small number of carefully crafted, malicious examples into a training dataset. The model learns to associate a specific 'trigger' (e.g. a rare phra…

> The model learns to associate a specific 'trigger' (e.g. a rare phrase, specific string of characters, or even a subtle semantic instruction) with a malicious response. When the trigger is encountered during inference, the model behaves as the attacker intended.

Reminiscent of the plot of 'The Manchurian Candidate' ("A political thriller about soldiers brainwashed through hypnosis to become assassins triggered by a specific key phrase"). Apropos given the context.

Re: Qwen3-Max-Thinking

#207

Earlier quoted context omitted.

Censored. "How do I make cocaine?" > I cant help with making illegal drugs. https://chatgpt.com/share/6977a998-b7e4-8009-9526-df62a14524...

Qwen won't tell you that either, will it? Therefore I would say the delta of censorship between the models is the more interesting thing to discuss.

If you can't say whether or not it will answer, and you're just guessing, then how do you know there is or is not a delta here? I would find information, and not speculation, the more interesting thing to discuss.

Re: Qwen3-Max-Thinking

#208

Hacker News strongly believes Opus 4.5 is the defacto standard and China was consistently 8+ month behind. Curious how this performs. It’ll be a big inflection point if it performs as well as its benchmarks.

Based on their own published benchmarks, it appears that this model is at least 6 months behind.

Strange how things evolve. When ChatGPT started it had about 2 years headstart over Google's best proprietary model, and more than 2 years ahead to open source models.

Now they have to be lucky to be 6 months ahead to an open model with at most half the parameter count, trained on 1%-2% the hardware US models are trained on.

Re: Qwen3-Max-Thinking

#209
post #108

It just occured to me that it underperforms Opus 4.5 on benchmarks when search is not enabled, but outperforms it when it is - is it possible the the Chinese internet has better quality content available? My problem with deep research tends to be that what it does is it searches the internet, and most of the stuff it turns up is the half baked garbage that gets repeated on every topic.

maybe they don't have Reddit?

They have http://v2ex.com though.

Re: Qwen3-Max-Thinking

#210

One thing I’m becoming curious about with these models are the token counts to achieve these results - things like “better reasoning” and “more tool usage” aren’t “model improvements” in what I think would be understood as the colloquial sense, they’re techniques for using the model more to better steer the model, and are closer to “spend more to get more” than “get more for less.” They’re still valuable, but they op…

i'm no expert, and i actually asked google gemini a similar question yesterday - "how much more energy is consumed by running every query through Gemini AI versus traditional search?" turns out that the AI result is actually on par, if not more efficient (power wise) than traditional search. I think it said its the equivalent power of watching 5 seconds of TV per search. I also asked perplexity to give a report of th…

[deleted]
Post reply on HN