Earlier quoted context omitted.
[flagged]
It's weird you got downvoted; you're correct, that chat bot was spewing hate speech at full blast, it was on the news everywhere. (For the uninformed: it didn't get unplugged for being "PR-unfriendly", it got unplugged because nearly every response turned into racism and misogyny in a matter of hours) https://en.wikipedia.org/wiki/Tay_(chatbot)#Initial_release
Qwen3-Max-Thinking
171–180 of 450 posts
Re: Qwen3-Max-Thinking
#172One thing I’m becoming curious about with these models are the token counts to achieve these results - things like “better reasoning” and “more tool usage” aren’t “model improvements” in what I think would be understood as the colloquial sense, they’re techniques for using the model more to better steer the model, and are closer to “spend more to get more” than “get more for less.” They’re still valuable, but they op…
I also asked perplexity to give a report of the most notable ARXIV papers. This one was at the top of the list -
"The most consequential intellectual development on arXiv is Sara Hooker's "On the Slow Death of Scaling," which systematically dismantles the decade-long consensus that computational scale drives progress. Hooker demonstrates that smaller models—Llama-3 8B and Aya 23 8B—now routinely outperform models with orders of magnitude more parameters, such as Falcon 180B and BLOOM 176B. This inversion suggests that the future of AI development will be determined not by raw compute, but by algorithmic innovations: instruction finetuning, model distillation, chain-of-thought reasoning, preference training, and retrieval-augmented generation. The implications are profound—progress is no longer the exclusive domain of well-capitalized labs, and academia can meaningfully compete again."
Re: Qwen3-Max-Thinking
#173I wasn't logged in so I don't have the ability to link to the conversation but I'm exporting it for my records.
Re: Qwen3-Max-Thinking
#174Re: Qwen3-Max-Thinking
#175[flagged]
As I recall reading in 2025, it has been proven that an actor can inject a small number of carefully crafted, malicious examples into a training dataset. The model learns to associate a specific 'trigger' (e.g. a rare phrase, specific string of characters, or even a subtle semantic instruction) with a malicious response. When the trigger is encountered during inference, the model behaves as the attacker intended.You can also directly modify a small number of model parameters to efficiently implement backdoors while preserving overall performance and still make the backdoor more difficult to detect through standard analysis. Further, can do tokenizer manipulation and modify the tokenizer files to cause unexpected behavior, such as inflating API costs, degrading service, or weakening safety filters, without altering the model weights themselves. Not saying any of that is being done here, but seems like a good place to have that discussion.
Re: Qwen3-Max-Thinking
#176Earlier quoted context omitted.
I've yet to encounter any censorship with Grok. Despite all the negative news about what people are telling it to do, I've found it very useful in discussing controversial topics. I'll use ChatGPT for other discussions but for highly-charged political topics, for example, Grok is the best for getting all sides of the argument no matter how offensive they might be.
Because something is offensive does not mean it reflects reality This reminds me of my classmates saying they watched Fox News “just so they could see both sides”
Re: Qwen3-Max-Thinking
#177Earlier quoted context omitted.
They've been quietly undoing a lot this IMO - gemini on the api will pretty much do anything other than CP.
Source? This would be pretty big news to the whole erotic roleplay community if true. Even just plain discussion, with no roleplay or fictional element whatsoever, of certain topics (obviously mature but otherwise wholesome ones, nothing abusive involved!) that's not strictly phrased to be extremely clinical and dehumanizing is straight-out rejected.
Would be very happy to see a source proving otherwise though; this has been a struggle to solve!
Re: Qwen3-Max-Thinking
#178[flagged]
Re: Qwen3-Max-Thinking
#179Earlier quoted context omitted.
That's not censorship, that's basic hygiene.
So you decide, then, how convenient for you.
Re: Qwen3-Max-Thinking
#180[flagged]
E.g. Qwen3 235B A22B Instruct 2507 gives an extensive reply starting with:
"The famous photograph you're referring to is commonly known as "Tank Man" or "The Tank Man of Tiananmen Square", an iconic image captured on June 5, 1989, in Beijing, China. In the photograph, a solitary man stands in front of a column of Type 59 tanks, blocking their path on a street east of Tiananmen Square. The tanks halt, and the man engages in a brief, tense exchange—climbing onto the tank, speaking to the crew—before being pulled away by bystanders. ..."
And later in the response even discusses the censorship:
"... In China, the event and the photograph are heavily censored. Access to the image or discussion of it is restricted through internet controls and state policy. This suppression has only increased its symbolic power globally—representing not just the act of protest, but also the ongoing struggle for free speech and historical truth. ..."