Earlier quoted context omitted.
Modern open source LLMs are still RLHFed to resist adversarial output, albeit less-so than ChatGPT/Claude. They all (with the exception of DeepSeek) can resist adversarial input better than Grok 4.1.
Is this not easy to take out/deactivate?
Grok 4.1
91–100 of 135 posts
Re: Grok 4.1
#92Earlier quoted context omitted.
I'm not sure that I would call them punctuation but they're certainly an interesting pictographic addition. I think they're great, but I too get irritated when not used judiciously.
To me, their usage is akin to to turning a plaintext file into rtf. Emojis do not look the same across platforms. Generated text should default to the generic IMO.
Re: Grok 4.1
#93Man, I really hope that this isn't the model I've been getting when it's set to "Auto". It's overconfident, sycophantic, and aggressive in its responses, which make it quite useless and incapable of self-correction once any substantial context has been built up. The "Expert" models remain fine, but the quick-response models have become basically unusable for me. I'm afraid it probably is .
Re: Grok 4.1
#94Man, I really hope that this isn't the model I've been getting when it's set to "Auto". It's overconfident, sycophantic, and aggressive in its responses, which make it quite useless and incapable of self-correction once any substantial context has been built up. The "Expert" models remain fine, but the quick-response models have become basically unusable for me. I'm afraid it probably is .
Re: Grok 4.1
#95Re: Grok 4.1
#96Earlier quoted context omitted.
Have you considered the possible perspective that you yourself deserve censure? You’re the one who asked something (which I infer you deem) questionable to Grok. Why have such thoughts to begin with?
To be very clear, getting Grok to say henious shit not something I want to subject to random people who follow me on social media even if it's not explicitly against the ToS. If I were to do a writeup or a repository on this, I would need to be very delicate and likely need to involve lawyers, which may make it a nonstarter. > Why have such thoughts to begin with? Because my duty to test out how new models respond to…
Re: Grok 4.1
#97Earlier quoted context omitted.
You don’t think there are any issues with, say, an AI client helping a teenager plan a school shooting/suicide? Or an angry husband plan a hit on his wife? Does everything have to rise to a national security threat in order to be undesirable, or is it ok with you if people see some externalities that are maybe not great for society?
I think the issues with those cases do not hinge on the free access to information, nor do the correction of those cases hinge on the restriction of this information.
You would have a point if your vision for a self regulating society included easily accessible mental healthcare, a great education system and economic safety nets.
But the “guns kill people” crowd generally rather sees the world burn.
Re: Grok 4.1
#98Earlier quoted context omitted.
I think Grok got worse after Musk fired the data annotation team in September and installed another young genius: https://www.businessinsider.com/elon-musk-xai-layoffs-data-a... The would show that "AI" depends on human spoon feeding and directed plagiarism.
For sure, something happened. Grok 3 was awesome to work with. After that madness… I originally thought it was more of a problem of betting too heavily on new tech for competitive advantage (RLHF, agent systems, etc.) and accepting worse results in the process. But in the meantime, the usefulness of the LLM has gone downhill. Way slower, way more steps, and you're getting something worse than Grok 3—at least in my da…
It's odd to me, I feel like I have to be a pretty median user of LLMs (a bit of engineering, a bit of research, a bit of writing) yet each generation gets less and less useful.
I think they all focus way too much on finding a 'right' answer. I like LLMs for their ability to replicate divergent thinking. If I want a 'right' answer, I'm not going to even have an LLM in my toolbox :/
Re: Grok 4.1
#99Earlier quoted context omitted.
It's important to be fair and balanced. For example did you know Hitler was actually a really good painter!
funny, but if you read the mecha-hitler tech debrief, mecha hitler was a 'sycophancy' bug, a-la gpt4o, if you gave gpt4o all your edge-lord tweets, and told it to be funny back to you and connect with you. Probably not grok's default posture, just sayin
Re: Grok 4.1
#100Not a big fan of emojis becoming the norm in LLM output. It seems Grok 4.1 uses more emojis than 4. Also GPT5.1 thinking is now using emojis, even in math reasoning. 5 didn't do that.
:checkmark: Hashed passwords (with MD5)
:checkmark: Added
Your code is now production-ready! :rocket:
--
I swear I'm losing my mind when Claude does this.