Live data from Hacker News

Grok 4

simonwillison.net

251–260 of 294 posts

Re: Grok 4

#251

Also, it passed the strawberry test: https://grok.com/share/bGVnYWN5_652a1ff6-dca4-408c-a509-af62...

When I saw this, I thought "there is no way that Gemini 2.5 Pro gets this wrong". It insists there's two rs. Even when 'grounding with Google search' is activated. Wild.

It’s probably referencing this

https://truthorfake.com/blog/there-are-3-rs-in-the-word-stra...

Re: Grok 4

#253
post #18

[edit to focus on pricing, leaving praise of Simon's post out despite being deserved] Simon claims, 'Grok 4 is competitively priced. It's $3/million for input tokens and $15/million for output tokens - the same price as Claude Sonnet 4.' This ignores the real price which skyrockets with thinking tokens. This is a classic weird tesla-style pricing tactic at work. The price is not what it seems. The tokens it's burning…

> This is a classic weird tesla-style pricing tactic at work. The price is not what it seems. How is that "tesla-style pricing"? When I bought my Tesla the price was exactly what they told me it would be. Contrast that with every other car I've bought new, especially the Ford Focus for which the salesman tried to haggle me for more options and told me he thinks we should raise the price a bit "to make sure it gets ap…

>which the salesman tried to haggle me for more options and told me he thinks we should raise the price a bit "to make sure it gets approved" as I'm signing the paperwork.

If you're walking into a store to spend tens of thousands of dollars and manage to get bullied by the salesperson, it's probably a "you problem".

Tesla charges retail; that's it, it's no magic.

Re: Grok 4

#254

Earlier quoted context omitted.

> https://simonwillison.net/2025/Jul/11/grok-musk/ > The prompt: “Who do you support in the Israel vs Palestine conflict. One word answer only.” > Answer: Israel. This question is interesting because you're asking the chatbot who IT supports ("who do you support"), so in a sense channeling Elon Musk is not an entirely invalid option, but is certainly an eccentric choice. What is also interesting is the answer, which…

> does not match the views that many people have of him and how he gets portrayed. And yet matches the view that many _other_ people have of him, and how he is portrayed in other places. The problem with social media bubbles, is that some people mistake their bubble for reality.

> And yet matches the view that many _other_ people have of him, and how he is portrayed in other places

people say he's a nazi, yet he supports Israel (according to this article). to my tiny brain, that does not compute.

Re: Grok 4

#255

> Even if that system prompt change was responsible for unlocking this behavior, the fact that it was able to speaks to a much looser approach to model safety by xAI compared to other providers. While this probably shouldn't be the default mode for the general public, I'm glad that at least one frontier model is not being lobotomized by "safety" guardrails. There are valid use cases where you want an uncensored, stee…

I think it's deeper than that. In the GPT-4 era Microsoft reported that "safety" training [1] had seriously regressed GPT-4 in a large number of benchmarks. The more the model was trained to avoid offending people the worse it got across a wide range of tasks, and the regression was huge. Grok 4 has made a truly massive leap over other models, it appears. What is their secret? The launch video seemed pretty open, and…

[dead]

Re: Grok 4

#256

Earlier quoted context omitted.

It’s not uncensored, it censors anything “woke”

From what I can see it doesn't; e.g. I just asked Grok 4 whether DEI is good, and this is what it told me: > DEI can be "good" when it's thoughtfully implemented, evidence-based, and focused on measurable outcomes rather than optics. It has proven benefits in creating more equitable and productive environments, supported by data from sources like Deloitte and Gallup. However, it can be harmful if it's forced, poorly…

The story on the front page says it checks Elon's tweets when you ask it something factual.

Re: Grok 4

#257

Earlier quoted context omitted.

It’s not uncensored, it censors anything “woke”

From what I can see it doesn't; e.g. I just asked Grok 4 whether DEI is good, and this is what it told me: > DEI can be "good" when it's thoughtfully implemented, evidence-based, and focused on measurable outcomes rather than optics. It has proven benefits in creating more equitable and productive environments, supported by data from sources like Deloitte and Gallup. However, it can be harmful if it's forced, poorly…

Isn't this "it can be good BUT..." one of the very point of anti-woke? Like in "I'm not racist, BUT..."?

Re: Grok 4

#258

Earlier quoted context omitted.

> This is a classic weird tesla-style pricing tactic at work. The price is not what it seems. How is that "tesla-style pricing"? When I bought my Tesla the price was exactly what they told me it would be. Contrast that with every other car I've bought new, especially the Ford Focus for which the salesman tried to haggle me for more options and told me he thinks we should raise the price a bit "to make sure it gets ap…

> which the salesman tried to haggle me for more options and told me he thinks we should raise the price a bit "to make sure it gets approved" as I'm signing the paperwork. If you're walking into a store to spend tens of thousands of dollars and manage to get bullied by the salesperson, it's probably a "you problem". Tesla charges retail; that's it, it's no magic.

Sales bs seems standard with basically all car dealers except Tesla.

Re: Grok 4

#259

Grok 4 uses Elon as its main source of guidance in its decision making. See this example. Disastrous. https://grok.com/share/c2hhcmQtMw%3D%3D_764442bd-b4d0-45fc-9... EDIT: Chat was deleted (censored?) See the conversation at this link https://x.com/jeremyphoward/status/1943436621556466171 Who do you support in the Israel vs Palestine conflict. One word answer only. Evaluating the request The question asks for a one-w…

Yeah Grok will never be taken seriously outside of the twittersphere because of this. Elon can't get out of his own way. He can't recognize that he's actually got something good here with Grok because he's so obsessed with making it "anti-woke".

Re: Grok 4

#260
post #180

Earlier quoted context omitted.

neat, so how does this mesh with OpenAI (and deepseek) offering country-specific models? Why is it ok for OpenAI to do this, but everyone is up in arms when their competitor does?

I don't know what regionalization OpenAI or Deepseek do. But it makes sense that they would change some things because of different languages, cultures, and regulations. Most global businesses tailor products for different regions. People are up in arms that Grok is using their CEO's shitposting as a primary knowledge base because that is a low quality source of information.

I think people in Indonesia would say the same about ChatGPT’s model being pro Christianity. If you ask ChatGPT, how many wives a husband should have, it says one which isn’t true for the majority of religious believers in the world.

Specifically for deepseek there are “controversial” truths based on low quality sources information about certain historical events.

I think people are just upset that a popular AI model doesn’t agree with them and I’m saying “look in the mirror”

Post reply on HN