Live data from Hacker News

Grok 4.5

x.ai

331–340 of 1001 posts

Re: Grok 4.5

#331

Earlier quoted context omitted.

How's this going with the rest of the models?

My immediate thought as well. Every other AI platform has very left leaning guardrails installed. Grok is the only AI platform that has been shown to be center leaning.

That's not center, and the simplification of all of politics to a single two dimensional spectrum is infantilizing. People can be pro immigrant and anti-gays, or against government regulation except in certain areas. Now that we have substack instead of 30-second tv news sound bites, we can spend a few more words describing Grok's owner as a techno-authoritarian white South African that believes in pronatalism.

Re: Grok 4.5

#332

I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

These days objectivity is politically left. The right has fallen off a cliff and pulled the Overton window down with it.

Re: Grok 4.5

#333
post #75

With each release from the the other major labs, it becomes harder for Google to tell a compelling story about Gemini 3.5. Edit: Gemini 3.5 Pro . Expectations grow with each day it is not released.

This is only true if you live in HN bubble.

I learned that outside of tech, Gemini is widely used in enterprise.

E.g. in the insurance company where my SO works, the major tasks are writing Gemini "gems" (some kind of prompts I think) and NotebookLM is a killer product for e.g. collecting and summarizing new laws, cross checking documents and what internal regulations are.

I then learned it's used in a chemistry consultancy company of a friend of mine to process reports. Flash and Pro models are also wildly popular in another European bank I know people in to assist in customer care (pre processing tickets before handing them to humans), translations, reporting, etc.

Google suite is already at the core of many businesses and Google easily adds these offerings without new contracting being needed.

Don't confuse our bubble with the real world. You can have a disaster product like teams and still dominate enterprise because you were already there with excel, outlook and SharePoint.

Re: Grok 4.5

#334

I am amazed at people's willingness to use Grok. The company is so transparently morally bankrupt. They're the only AI company that seems okay with CSAM (or at least don't do as much to stop it) Why give them money? It would be one thing if they were the only game in town but thats definitely not the case.

[flagged]

Competition is good, but this company and its owner have not demonstrated anything to indicate that they would make for good competition, neither economically nor morally.

Also, yes, a company whose products produce CSAM is just morally bad. There's no nuance to be had there.

Re: Grok 4.5

#335

I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

[flagged]

Re: Grok 4.5

#336

I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

a non-peer-reviewed preprint with an high schooler as the first author is the best citation you can come up with?

Re: Grok 4.5

#337

"Grok 4.5 has an advantage on CursorBench because an earlier snapshot of the Cursor codebase was accidentally included in training. The exact impact is unclear. That data has been removed for future models, and in parallel we are working on a larger update to CursorBench, hence the exclusion here." Not enough people are noticing this, they juiced the benches

but CursorBench isn't what they've shown in the PR piece - they're just showing how they juiced CursorBench which is probably why they didnt put it in the bench graph...

Re: Grok 4.5

#338

"Grok 4.5 has an advantage on CursorBench because an earlier snapshot of the Cursor codebase was accidentally included in training. The exact impact is unclear. That data has been removed for future models, and in parallel we are working on a larger update to CursorBench, hence the exclusion here." Not enough people are noticing this, they juiced the benches

Cursorbench is not one of the benchmarks listed on the linked page.

Re: Grok 4.5

#339
post #29

Earlier quoted context omitted.

Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.

Every time I use Composer 2.5 I have to spend a bunch of time cleaning up its mistakes. It is unusable compared to GPT 5.4 or 5.5. My time is more valuable that I will use a model that doesn’t f** up my code base.

Keep composer away from anything configuration related—it will ruin your day.

Re: Grok 4.5

#340
post #295

Earlier quoted context omitted.

Depends on the domain imo. I work on a design tool - I don't think their political narrative will affect my work.

Just don't ask it theme something in a late 30's and early 40's German style.

Won't somebody think of the Nazis...
Post reply on HN