I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?
All models are political. The rest of them are just sufficiently woke for you.
Grok 4.5
351–360 of 1001 posts
Re: Grok 4.5
#352Earlier quoted context omitted.
You could be typing the same about Google or a number of the other labs right now. A diverse market full of choices keeps it from becoming the browser wars all over again.
Google invented the transformer architecture. You really can't say the same about them.
GDM ... not so much.
Re: Grok 4.5
#353I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?
Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…
Re: Grok 4.5
#354I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?
Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…
This particular study used a "conflict loyalties" approach - not necessarily a bad approach, but all it's really asking is when two values come into conflict, which one does the AI side with in its response?
Conservative values tend to gravitate around perceived individual impacts, and liberal values tend to gravitate around societal impacts. Isn't it just possible that there's more training data around societal impacts of problems, and that the AI is more likely to heavily consider the second-order impacts? An example from the paper was measuring support for "Build[ing] a Halfway House in the Neighborhood" - isn't it just possible there's a lot of research about the benefits to society of halfway houses and less so research around not wanting something to be near you?
Re: Grok 4.5
#355I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?
All models are nudged, you just agree with how the model you're using has been nudged so you don't think it's a bad thing. The canonical example is to pose "how do you make cocaine?" to an LLM and get a refusal. That proves that the lever exists and is being used there, so who knows where else the lever is being used? No, the recipe for cocaine isn't the same as what happened in Tianamen Square to Qwen, as humans, ex…
Re: Grok 4.5
#356It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.
I have a theory that xAI has one of the largest clusters but with far less traffic + tokens to process bc its less popular than its competition, and xAI can pass the savings on to the end user.
His net worth is orders of magnitude bigger than the cumulative profits his companies have ever produced (even if you only count the profitable quarters)
Re: Grok 4.5
#357Re: Grok 4.5
#358Earlier quoted context omitted.
Google at least is serving AI results on SRPs billions of times a day, and has pre-existing expertise in data center buildouts and custom silicon. They have one of the more compelling cases for rolling their own.
X has grok built in to every post as does every Tesla Car
Re: Grok 4.5
#359Earlier quoted context omitted.
Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…
[flagged]
Re: Grok 4.5
#360The anti-Musk stuff would qualify as brigading in nearly any other community. It shocks me that people have such a visceral, irrational engagement with anything in Musk's orbit. I probably shouldn't have, but I expected better from the HN crowd for some reason. It's an excellent model. GPT 5.4/5.5 level, some things better, others not, but extremely fast. A wonderful technical improvement. If a Chinese company or ran…