Earlier quoted context omitted.
How's this going with the rest of the models?
My immediate thought as well. Every other AI platform has very left leaning guardrails installed. Grok is the only AI platform that has been shown to be center leaning.
Grok 4.5
331–340 of 1001 posts
Re: Grok 4.5
#332I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?
Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…
Re: Grok 4.5
#333With each release from the the other major labs, it becomes harder for Google to tell a compelling story about Gemini 3.5. Edit: Gemini 3.5 Pro . Expectations grow with each day it is not released.
I learned that outside of tech, Gemini is widely used in enterprise.
E.g. in the insurance company where my SO works, the major tasks are writing Gemini "gems" (some kind of prompts I think) and NotebookLM is a killer product for e.g. collecting and summarizing new laws, cross checking documents and what internal regulations are.
I then learned it's used in a chemistry consultancy company of a friend of mine to process reports. Flash and Pro models are also wildly popular in another European bank I know people in to assist in customer care (pre processing tickets before handing them to humans), translations, reporting, etc.
Google suite is already at the core of many businesses and Google easily adds these offerings without new contracting being needed.
Don't confuse our bubble with the real world. You can have a disaster product like teams and still dominate enterprise because you were already there with excel, outlook and SharePoint.
Re: Grok 4.5
#334I am amazed at people's willingness to use Grok. The company is so transparently morally bankrupt. They're the only AI company that seems okay with CSAM (or at least don't do as much to stop it) Why give them money? It would be one thing if they were the only game in town but thats definitely not the case.
[flagged]
Also, yes, a company whose products produce CSAM is just morally bad. There's no nuance to be had there.
Re: Grok 4.5
#335I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?
Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…
Re: Grok 4.5
#336I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowledge that their models are being nudged around in the backend?
Studies on political bias in models consistently show that LLMs lean politically left. The only outlier is grok which leans right but by a smaller factor, according to this study for instance: https://arxiv.org/abs/2603.23841 Edit: adding some other studies that are easily retrievable with a quick search for those unsatisfied with the first one - https://arxiv.org/abs/2606.12922 https://arxiv.org/abs/2412.16746 Claim…
Re: Grok 4.5
#337"Grok 4.5 has an advantage on CursorBench because an earlier snapshot of the Cursor codebase was accidentally included in training. The exact impact is unclear. That data has been removed for future models, and in parallel we are working on a larger update to CursorBench, hence the exclusion here." Not enough people are noticing this, they juiced the benches
Re: Grok 4.5
#338"Grok 4.5 has an advantage on CursorBench because an earlier snapshot of the Cursor codebase was accidentally included in training. The exact impact is unclear. That data has been removed for future models, and in parallel we are working on a larger update to CursorBench, hence the exclusion here." Not enough people are noticing this, they juiced the benches
Re: Grok 4.5
#339Earlier quoted context omitted.
Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.
Every time I use Composer 2.5 I have to spend a bunch of time cleaning up its mistakes. It is unusable compared to GPT 5.4 or 5.5. My time is more valuable that I will use a model that doesn’t f** up my code base.