It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.
Grok 4.5 is a huge step up from their next best model and now around the same performance as GLM 5.2, but it's not exactly at the frontier of the cost efficiency curve in our coding evaluations. That curve is defined by the 2 lighter GPT 5.6 models. However, the fact that they finally have a strong post-training and RL setup bodes well for future releases. They certainly are not compute-constrained anymore. Data at h…
Grok 4.5
991–1000 of 1001 posts
Re: Grok 4.5
#992Earlier quoted context omitted.
You could be typing the same about Google or a number of the other labs right now. A diverse market full of choices keeps it from becoming the browser wars all over again.
> A diverse market full of choices keeps it from becoming the browser wars all over again. This is a great analogy but I worry you might be implying something I don't agree with but you didn't explicitly say what I'm worried about, so let me call it out: Microsoft played a dirty game with I.E, but they are in the dirty game business. It wasn't only I.E, it was their OS, Office suite and everything else they do busine…
There used to be 4+ major browser engines. Now there's only two, and Google owns almost complete share & following standards? Google prefers their monoculture.
More diversity in browser engines was a good thing & standards were lovely.
Not everything has to be a slug-fest between #1 & #2.
& personally I'm glad there's Grok & Gemini to keep Anthropic & OpenAI on their toes even more than the competitive band of open weights models already do.
The corporate models tend to be defensive and sound like HR has approved every word of the script. Not a fan.
Re: Grok 4.5
#993Earlier quoted context omitted.
What a crazy thing to say explicitly about the model that _avoids_ that kind of stuff. Did we already collectively forget about the ChatGPT "misgendering is worse than thermonuclear war" case, or Gemini picturing the Founding Fathers as black? Anything political, I always go to Grok first. It's the only one that has bled for not just trying to play it super safe with political correctness, but trying to be impartial.
You forget about Mecha-Hitler and Musk having the system prompts changed for him? Grok is dead-last on impartiality.
I want a car that If I decide to run it in to a wall to avoid running over people, it does that. People arguing that Grok doing what the user asks for, being actually impartial, being somehow "biased" is crazy.
Re: Grok 4.5
#994Earlier quoted context omitted.
What a crazy thing to say explicitly about the model that _avoids_ that kind of stuff. Did we already collectively forget about the ChatGPT "misgendering is worse than thermonuclear war" case, or Gemini picturing the Founding Fathers as black? Anything political, I always go to Grok first. It's the only one that has bled for not just trying to play it super safe with political correctness, but trying to be impartial.
>Anything political, I always go to Grok first. Why would you talk to an LLM about anything political?
Re: Grok 4.5
#995I echo the top comment - the politics here have gotten out of control. Like it or not, Elon and his companies have changed the world for the better. You are allowed to separate the human from the human's impact sometimes.
I kind of doubt that Twitter/X is any sort of net benefit. Boring company was a wash, Tesla is a net positive, although it also has some issues. Paypal was a huge advancement, although it's still deeply flawed. Grok may end up important, but right now it's kind of just hovering around 'acceptable' second-tier. But SpaceX has the potential of driving some of the grandest and most revolutionary accomplishments of the 2…
Re: Grok 4.5
#996So depressing to read the non-stop political comments here. I’m using GLM at the moment, it’s Chinese and backed by god knows who, and no one cares. I really wanted to see what the experts thought of new Grok tech and how the model compares etc. I wish I could turn off the non-technical comments somehow, could literally just go to reddit if I want to see garbage like this. Am I supposed to get emotional every time I…
Thank you. I too am interested in the technical aspects, not how many guardrails it can hit against Fable, like, who cares? As long as it can get me (more) uncensored access without killing itself over biology or chemistry, it's a good model to me. I can't believe the top comment is about some political reply garbage, as if that actually matters day to day for coders; in reality I want to get my work done.
Re: Grok 4.5
#997Earlier quoted context omitted.
Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.
It is hard to evaluate the model performance of Composer 2.5 when Cursor's harness is so awful compared to the others on the market.
Re: Grok 4.5
#998Earlier quoted context omitted.
I have a theory that xAI has one of the largest clusters but with far less traffic + tokens to process bc its less popular than its competition, and xAI can pass the savings on to the end user.
xAI had $2.5B in operating losses in the past quarter. What savings are being passed on?
Re: Grok 4.5
#999Earlier quoted context omitted.
Someone has to know. Would be nice if an insider would drop some hints so that the open-source space could make some good progress.
Nobody has to actually know the secret of their own success, especially not relative success to equally-secretive near-peers. Same as with rich person autobiographies: even when they tell you what they think it is, they can't see the path not travelled.
Yup, there's a lot of survivorship bias in those. And humans want to attribute success to some skill somehow. You cannot just have been lucky.
Re: Grok 4.5
#1000Reading through these threads is cursed. This site supposedly attracts the brightest of us but its all just a bunch of people ignoring the article and screaming that Musk is evil so everything he does must be ignored and villified. Just really dissapointing hysteria on display here.
That's not a very high bar today.