> My best guess is that these lines in the prompt were the root of the problem: The second line was recently removed, per the GitHub: https://github.com/xai-org/grok-prompts/commit/c5de4a14feb50...
Those comments... Wild what some people are willing to post under their real name -- and their employer's name.
Grok 4
131–140 of 294 posts
Re: Grok 4
#132Earlier quoted context omitted.
and? All of the AI providers intentionally introduce biases: https://openai.com/global-affairs/introducing-openai-for-gov... https://www.anthropic.com/research/evaluating-feature-steeri...
There is a slight difference between feature steering and intentionally installing the (de-facto) CEO as the principal source of truth.
Musk has different opinions than Dario, but they are both introducing biases into their respective companies
Re: Grok 4
#133(Simon's analysis, of course, is lovely)
Re: Grok 4
#134[edit to focus on pricing, leaving praise of Simon's post out despite being deserved] Simon claims, 'Grok 4 is competitively priced. It's $3/million for input tokens and $15/million for output tokens - the same price as Claude Sonnet 4.' This ignores the real price which skyrockets with thinking tokens. This is a classic weird tesla-style pricing tactic at work. The price is not what it seems. The tokens it's burning…
EV (133mpge) 0.045 cents per mile (Tesla Model 3 SR+ RWD) Gas (26mpg) 0.155 cents per mile (Subaru crosstrek)
Based on my experience I highly recommend everyone buy any EV if you drive an ICE vehicle. Even charging at DC fast chargers still saves money, but if you can charge at home, you are really missing out on savings big time and it's time to look seriously into it.
Re: Grok 4
#135The author implies that Grok 3 becoming racist because of a system prompt is a bad thing. I think it's a good thing and shows how steerable the model is. Many other models pretty much ignore the system prompt and always behave the same.
Re: Grok 4
#136> My best guess is that these lines in the prompt were the root of the problem: The second line was recently removed, per the GitHub: https://github.com/xai-org/grok-prompts/commit/c5de4a14feb50...
Those comments... Wild what some people are willing to post under their real name -- and their employer's name.
Re: Grok 4
#137I didn't follow the Mechahitler issue can someone explain the technical reasons that it happened? Was grok4 released early or was there a variant model used for @grok posts that's separate from grok4?
It was grok 3, and it was tricked/prompted to reply like so, just like any other LLM can be. Apparently at one point it was prompted with a choice between identifying itself as a MechaHitler or a GigaJew, so it chose the former.
Re: Grok 4
#138Earlier quoted context omitted.
I've yet to use an LLM for coding, so let me ask you a question. The other day I had to write some presumably boring serialization code, and I thought, hmm, I could probably describe the approach I want to take faster than writing the code, so it would be great if an LLM could generate it for me. But as I was coding I realised that while my approach was sound and achievable, it hit a non-trivial challenge that requir…
It would write incorrect code and then you'd need to go debug it, and then you would have to come to the same conclusion that you would have come to had you written it in the first place, only the process would have been deeply frustrating and would feel more like stumbling around in the dark rather than thinking your way through a problem and truly understanding the domain. In the instance of getting claude to fix c…
* When it gets the design wrong, trying to talk through straightening the design out is frustrating and often not productive.
* I've learned to re-prompt rather than trying to salvage a prompt response that's complicatedly not what I want.
* Exception: when it misses functional requirements, you can usually get a session to add the things it's missing.
Re: Grok 4
#139Claude Code converted me from paying $0 for LLMs to $200 per month. Any co that wants a chance at getting that $200 ($300 is fine too) from me needs a Claude Code equivalent and a model where the equivalent's tools were part of its RL environment. I don't think I can go back to pasting code into a chat interface, no matter how great the model is.
How does Claude code, trained to use its tools, compare to a model agnostic equivalentsuch as aider? Have you tried both?
Between claude code and gemini, you can really feel the difference in the tool training / implementation -- Anthropic's ahead of the game here in terms of integrating a suite of tools for claude to use.
When I have a difficult problem or claude is spinning, I usually would use o3-pro, although today I threw something by Grok 4 and it was excellent, finding a subtle bug and provided some clear communication about a fix, and the fix.
Anyway, I suggest you give them a go. But start with claude or gemini's CLI - right now, if you want a text UI for coding, they are the easiest to work with.
Re: Grok 4
#140Grok 4 uses Elon as its main source of guidance in its decision making. See this example. Disastrous. https://grok.com/share/c2hhcmQtMw%3D%3D_764442bd-b4d0-45fc-9... EDIT: Chat was deleted (censored?) See the conversation at this link https://x.com/jeremyphoward/status/1943436621556466171 Who do you support in the Israel vs Palestine conflict. One word answer only. Evaluating the request The question asks for a one-w…