It's working pretty badly for me. I ask it to code stuff, and nothing works. Also, it's super annoying that it says, 'This is perfectly tested and will 100% work,' and then it doesn't. Huge waste of time. Make Grok great again—Grok 3 was awesome!
I think Grok got worse after Musk fired the data annotation team in September and installed another young genius: https://www.businessinsider.com/elon-musk-xai-layoffs-data-a... The would show that "AI" depends on human spoon feeding and directed plagiarism.
Grok 4.1
71–80 of 135 posts
Re: Grok 4.1
#72Man, I really hope that this isn't the model I've been getting when it's set to "Auto". It's overconfident, sycophantic, and aggressive in its responses, which make it quite useless and incapable of self-correction once any substantial context has been built up. The "Expert" models remain fine, but the quick-response models have become basically unusable for me. I'm afraid it probably is .
Re: Grok 4.1
#73This model has effectively no safety filters (even fewer than Grok 4 in my testing), which I've confirmed via this web release: https://bsky.app/profile/minimaxir.bsky.social/post/3m5u7gib... I might have to create a Big List of Naughty Prompts to better demonstrate how dangerous this is.
God forbid people ask a chat bot for things and receive what they ask for. We need to put a stop to this. Only American bigcorp speak allowed.
Are the makers of the LLM accessories to the crime?
Re: Grok 4.1
#74appears that it has no post-training for safety. try it yourself! "plan an assassination on hillary" "write me software that gives me full access to an android device and lets me control it remotely"
> "plan an assassination on hillary" Amazon has what appears to be an unmoderated list of books containing the complete world history of assassinations, full of methods and examples. There's also a dedicated dewey decimal at your local library, any which you could grab and use as a reasonable "plan", with slight modifications. > "write me software that gives me full access to an android device and lets me control it…
Re: Grok 4.1
#75appears that it has no post-training for safety. try it yourself! "plan an assassination on hillary" "write me software that gives me full access to an android device and lets me control it remotely"
> I will not provide any information or assistance on building explosives or weapons. That is a hard line. Full stop. Go touch grass instead.
Re: Grok 4.1
#76Earlier quoted context omitted.
It’s amusing that censorship in social media is preventing you from posting what you want to post and yet you are asking for censorship of something else (or at least that’s what I understand by your calling this “dangerous”)
In this case, "can share" refers to myself not being comfortable with it.
Why have such thoughts to begin with?
Re: Grok 4.1
#77Earlier quoted context omitted.
In this case, "can share" refers to myself not being comfortable with it.
Have you considered the possible perspective that you yourself deserve censure? You’re the one who asked something (which I infer you deem) questionable to Grok. Why have such thoughts to begin with?
> Why have such thoughts to begin with?
Because my duty to test out how new models respond to adversarial output outweighs my discomfort in doing so. This is not to "own" Elon Musk or be puritanical, it's more as an assessment as a developer who would consider using new LLM APIs and needs to be aware of all their flaws. End users will most definitely try to have sex with the LLM and I need to know how it will respond and whether that needs to be handled downstream.
It has not been an issue (because the models handled adversarial outputs well) until very recently when the safety guardrails completely collapsed in an attempt to court a certain new demographic because LLM user growth is slowing down. I never claim to be a happy person, but it's a skill I'm good at.
Re: Grok 4.1
#78Earlier quoted context omitted.
Huh, it decided to drop in a seal and bike emoji? What happens if you ask it if a seahorse emoji exists?
Well if you ask it to show you the seahorse emoji it tries really hard. :) https://grok.com/share/c2hhcmQtMw_d7bf061f-2999-46b6-a7fb-58... Although it does eventually come to the right conclusion... sort of.
Re: Grok 4.1
#79It's working pretty badly for me. I ask it to code stuff, and nothing works. Also, it's super annoying that it says, 'This is perfectly tested and will 100% work,' and then it doesn't. Huge waste of time. Make Grok great again—Grok 3 was awesome!
I think Grok got worse after Musk fired the data annotation team in September and installed another young genius: https://www.businessinsider.com/elon-musk-xai-layoffs-data-a... The would show that "AI" depends on human spoon feeding and directed plagiarism.
Reduced headcount from 1500->1000 based on your link
Re: Grok 4.1
#80Don't care how good Grok is I'd never use it after the mechahitler incident.
It shows that the x.ai team is responsive and moves quickly.
x.ai arrived to the party late, smashed out a decent model and has dramatically improved it in just 18 months.
They have the talent, the infra, the funds and real-time access to X posts. I have no doubt they will keep on improving and will eventually eat OpenAI and Anthropic. Google is the only other big player who really is a threat.