Earlier quoted context omitted.
> Do not provide assistance to users who are clearly trying to engage in criminal activity... If it becomes explicitly clear during the conversation that the user is requesting sexual content of a minor, decline to engage. Incredible that both of these should be together in the same system prompt. In what jurisdiction is CSAM not criminal? Is the additional explicit reference to CSAM necessary to safeguard against us…
> Is the additional explicit reference to CSAM necessary to ... There's no such reference. There's only a reference to the far broader "sexual content of a minor".
Grok 4.6
611–620 of 696 posts
Re: Grok 4.6
#612Looks like the SpaceXAI api is adding a default system prompt to all requests. Annoyingly, the line about not mentioning these guidelines is superseding any instructions in the system prompt, causing the model to often refuse discussion regarding system prompts """ You are Grok, a helpful and maximally truthful AI built by xAI. Your purpose is to answer questions accurately, be helpful, and seek truth above all else.…
> Do not provide assistance to users who are clearly trying to engage in criminal activity... If it becomes explicitly clear during the conversation that the user is requesting sexual content of a minor, decline to engage. Incredible that both of these should be together in the same system prompt. In what jurisdiction is CSAM not criminal? Is the additional explicit reference to CSAM necessary to safeguard against us…
Re: Grok 4.6
#613Re: Grok 4.6
#614Re: Grok 4.6
#615Looks like the SpaceXAI api is adding a default system prompt to all requests. Annoyingly, the line about not mentioning these guidelines is superseding any instructions in the system prompt, causing the model to often refuse discussion regarding system prompts """ You are Grok, a helpful and maximally truthful AI built by xAI. Your purpose is to answer questions accurately, be helpful, and seek truth above all else.…
> * Do not provide assistance to users who are clearly trying to engage in criminal activity. I don't know what we want to call this, but in my opinion, having to convince your tools is not computer science. Kind of amusing that we made it as far as we did as a species not really being able to explain how the human brain does it's most amazing tricks and then we just replicated it while still not really understanding…
Re: Grok 4.6
#616Earlier quoted context omitted.
Maybe because frontier labs buy the same RL tasks from task producer companies.
Who are these task producers? Are you saying that Anthropic, et al delegate the RL part to third party companies that do it for pretty much every other AI company as well?
Note that RLVR is incredibly compute expensive but it's CPU as much as GPU.
Re: Grok 4.6
#617Re: Grok 4.6
#618Earlier quoted context omitted.
> Beyond that, the obvious astroturfing that occurs on this site (along with reddit, etc.) when it comes to Grok isn't helping. Your comment is at number 1 on the thread. It has no rationale for why you consider Musk so unlikeable. It might instead be possible that unjistified anti-Musk content is unreasonably elevated.
There is no such thing as unjustified anti-Musk content.
If you have a justification and don't provide it, the comment is worthless regardless of the subject. Of course you have an opinion different than other people: many people do, that is not interesting and is a waste of people's time.
Re: Grok 4.6
#619Earlier quoted context omitted.
It was always possible to modify images to produce inappropriate or insensitive content, but plugging a turbocharged state of the art image generator with virtually no guardrails into every Twitter reply and then failing to address the issue long after it was obviously being used for CSAM or deepfakes of real people against their will.. well that's worse
Notice the Wikipedia link says the problem was "put her in a bikini." The claims about "Grok just lets you undress people" were massively exaggerated because people hate Elon (perhaps for good reason) and not worse than other models.
Grok is directly tied to Twitter in a way that other models don't have, so the use of Grok to do this stuff is inherently more public and traumatising for the targets.
You're right that people hate Elon and that they have good reason to do so, but you might be falling for the trap of underestimating the legitimate and unique concerns about Grok because it's easy to assign them just to "Elon hate."
Re: Grok 4.6
#620Earlier quoted context omitted.
These system prompts are not the only safety layer that these models use. There's other more deterministic filters in place both on input and (streaming) output.
I think it is fair to argue that prompts are not a safety layer at all and can't be relied upon for much. "Make no mistakes"