Live data from Hacker News

Grok 4.5

x.ai

351–360 of 1001 posts

Re: Grok 4.5

#351

I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?

All models are political. The rest of them are just sufficiently woke for you.

Grok was ranting about "white genocide" in unrelated conversations just one year ago.

Re: Grok 4.5

#352

Earlier quoted context omitted.

You could be typing the same about Google or a number of the other labs right now. A diverse market full of choices keeps it from becoming the browser wars all over again.

Google invented the transformer architecture. You really can't say the same about them.

Google Brain invented the transformer.

GDM ... not so much.

Re: Grok 4.5

#353

I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

So tired of this false equivalence between left and right politics in the US. So because major LLMs support a increase in the minimum wage we need to offset it with an LLM that has spouted Nazi propaganda and lets you generate porn image of real people?

Re: Grok 4.5

#354

I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

You do realize that the scale on papers like this is the important part, and the creation of that scale is itself inherently biased?

This particular study used a "conflict loyalties" approach - not necessarily a bad approach, but all it's really asking is when two values come into conflict, which one does the AI side with in its response?

Conservative values tend to gravitate around perceived individual impacts, and liberal values tend to gravitate around societal impacts. Isn't it just possible that there's more training data around societal impacts of problems, and that the AI is more likely to heavily consider the second-order impacts? An example from the paper was measuring support for "Build[ing] a Halfway House in the Neighborhood" - isn't it just possible there's a lot of research about the benefits to society of halfway houses and less so research around not wanting something to be near you?

Re: Grok 4.5

#355

I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?

All models are nudged, you just agree with how the model you're using has been nudged so you don't think it's a bad thing. The canonical example is to pose "how do you make cocaine?" to an LLM and get a refusal. That proves that the lever exists and is being used there, so who knows where else the lever is being used? No, the recipe for cocaine isn't the same as what happened in Tianamen Square to Qwen, as humans, ex…

[deleted]

Re: Grok 4.5

#356
post #5

It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.

I have a theory that xAI has one of the largest clusters but with far less traffic + tokens to process bc its less popular than its competition, and xAI can pass the savings on to the end user.

Profitability is never a constraint for Elon companies. He has always been able to be able to extract money from the middle east, government, banks, retail investors (or these same parties through his other companies) whenever they need more.

His net worth is orders of magnitude bigger than the cumulative profits his companies have ever produced (even if you only count the profitable quarters)

Re: Grok 4.5

#358
post #144

Earlier quoted context omitted.

Google at least is serving AI results on SRPs billions of times a day, and has pre-existing expertise in data center buildouts and custom silicon. They have one of the more compelling cases for rolling their own.

X has grok built in to every post as does every Tesla Car

/s

Re: Grok 4.5

#359

Earlier quoted context omitted.

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

[flagged]

Unless you ask them about IQ, which is apparently not real?

Re: Grok 4.5

#360

The anti-Musk stuff would qualify as brigading in nearly any other community. It shocks me that people have such a visceral, irrational engagement with anything in Musk's orbit. I probably shouldn't have, but I expected better from the HN crowd for some reason. It's an excellent model. GPT 5.4/5.5 level, some things better, others not, but extremely fast. A wonderful technical improvement. If a Chinese company or ran…

[flagged]
Post reply on HN