Live data from Hacker News

Grok 4.5

x.ai

371–380 of 1001 posts

Re: Grok 4.5

#371

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

> Can someone breakdown to me how this makes any sort of economical sense?

Capital markets are excited by AI.

By tying his rockets to AI with his vision of “orbital data centers”, Elon turned an $8 per share IPO (at least according to financial times and morgan stanley) into a $135 per share (1.8T) IPO.

Re: Grok 4.5

#372

"Grok 4.5 has an advantage on CursorBench because an earlier snapshot of the Cursor codebase was accidentally included in training. The exact impact is unclear. That data has been removed for future models, and in parallel we are working on a larger update to CursorBench, hence the exclusion here." Not enough people are noticing this, they juiced the benches

They're saying they didn't include the benchmark which errantly leaked into the training data.

Re: Grok 4.5

#373

Earlier quoted context omitted.

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

[flagged]

Correction: facts are facts.

How a person perceives facts categorizes them into a political bias bucket.

Re: Grok 4.5

#375

Earlier quoted context omitted.

Google is playing a different game. I don't really know what game they're playing, but they're not trying to beat Claude Code. They have coding capabilities and Antigravity, but I'd be surprised if it's much more than an afterthought. They're focusing on efficiency, models at the edge, human interaction, image and video, etc. in ways Anthropic, in particular, is not. Google wants its AI to be pervasive in everyone's…

> Google wants its AI to be pervasive in everyone's daily life. Google wants nothing more than the world to remain stuck in 2000 - 2020 where search was king. Their organisational inertia will fight its AI progress every step of the way and this very well explains why they are not leading the AI pack despite inventing the technology.

I mean, I'm sure there are people within Google who are behaving as though they can keep the dream of the 00s alive in Mountain View, but there's also a whole bunch of people doing work at the frontier in AI. Google has a large lead in hardware, they have the smartest very small models, they have among the most efficient large models (I'd wager their margins on Gemini 3.5 Flash inference are absurd). They have among the best image, video, and audio models, going in every direction (generating, editing, understanding).

Viewed from a consumer lens, the AI the average person interacts with daily, Google seems like the clear leader, especially after locking in Apple as a customer for iPhones.

Re: Grok 4.5

#376

Earlier quoted context omitted.

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

Reality leans left in many respects, principally the non-economic ones. It's a simple consequence of the same trend of overall social and educational progress that allowed these models to be developed in the first place. There is a reason they came out of San Francisco, and not Russia, Iran, or Oklahoma. To get a right-biased response from an LLM, you have to deliberately bias it... which is exactly what Musk did. Ne…

This fake intellectual nonsense is exactly why there should be deep institutional scrutiny of sota models. Elon is the worst person to do this but he's 'not wrong' that there needs to be scrutiny

Re: Grok 4.5

#377

Earlier quoted context omitted.

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

Models are tuned to give ethical responses and right-leaning responses are judged by RLHF process to be unethical. That's your problem. It's not 'left' or 'right' to be ethical, but if one side is inherently antisocial and unethical then it's going to naturally create an appearance of bias toward the other.

Who determines what is "ethical"?

Re: Grok 4.5

#378

I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?

By not using them on something political? Why do I care when I'll just use it to generate code?

Because a percentage of every dollar you spend on it will go towards pushing political opinions that run contrary to your own best interests?

Re: Grok 4.5

#379

Earlier quoted context omitted.

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

Models have a goal of accuracy and accuracy is not the median of the left/right spectrum.

[deleted]

Re: Grok 4.5

#380

Earlier quoted context omitted.

Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…

Models are tuned to give ethical responses and right-leaning responses are judged by RLHF process to be unethical. That's your problem. It's not 'left' or 'right' to be ethical, but if one side is inherently antisocial and unethical then it's going to naturally create an appearance of bias toward the other.

Yeah nobody wants to balance their answers with Nazi ideology (far right basically) in their tuning except Grok
Post reply on HN