Live data from Hacker News

Grok 4.5

x.ai

271–280 of 1001 posts

Re: Grok 4.5

#272
post #222

Earlier quoted context omitted.

I have a theory that xAI has one of the largest clusters but with far less traffic + tokens to process bc its less popular than its competition, and xAI can pass the savings on to the end user.

Why would having more costs and less income allow them to pass savings on to the end user?

“We lose money on every rack, but we make up for it in volume!” - Elon Musk, probably

Re: Grok 4.5

#273
post #5

It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.

I have a theory that xAI has one of the largest clusters but with far less traffic + tokens to process bc its less popular than its competition, and xAI can pass the savings on to the end user.

they are renting parts to google for like 1b a month

really dont think they have a lot of idle power

Re: Grok 4.5

#274
post #29

Earlier quoted context omitted.

Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.

It is hard to evaluate the model performance of Composer 2.5 when Cursor's harness is so awful compared to the others on the market.

Not my experience at all. I've been using Cursor hardcore for about two weeks now and Composer 2.5 and it's wonderful. Now with Grok 4.5 I'm quite excited about the possibilities.

Re: Grok 4.5

#275

(from Cursor's blog) > Training included trillions of tokens of Cursor data which capture a wide-range of user interactions with codebases and software tools. This dataset lets the model learn both from existing software as well as developer-agent interactions, capturing how developers work and how agents interact with their environments. This is what the big money was for. Cursor is the first big player that had rea…

well the big money was also in spacex stock, fresh post IPO, so overall a very smart move it seems

Re: Grok 4.5

#276

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

Anthropic is already profitable, economics is no longer an issue as they have found PMF in enterprise software market. You might need to update your views.

https://www.wsj.com/tech/ai/mind-blowing-growth-is-about-to-...

Re: Grok 4.5

#277

Earlier quoted context omitted.

I think they have a better agent personality which pushes back and isn't sycophantic. It has been awhile since I've used the others but that's where it locked me in and I've stuck with it.

> isn't sycophantic Not sure about that one... But I think the true secret sauce for all these models is how they reason. GPT never outputs how it thinks, which "saves on tokens" but Claude absolutely tells you how it thinks, and there's people who use how it reasons about solving problems to finetune smaller open source models, with surprisingly better output.

From my experience, it has not been sycophantic in the sense that it pushes back and questions my own reasoning in healthy ways. There were moments where I felt I was brushing up against actual AI psychosis, and it pushed back on my questioning of its intentions, that it even had intentions. I'll put it this way: I feel comfortable recommending Claude to people who haven't experienced AI yet. As we've learned from early experiences with other models, leading people down paths of believing they understood math in ways nobody else has and even harming themselves, I put Claude as a safer alternative.

Re: Grok 4.5

#278

Earlier quoted context omitted.

Grok is the #1 uncensored easily-available model, and it's also tightly integrated with Twitter.

Is uncensored a selling point? What do people use uncensored Grok for (like, real use cases) that they can't or won't use other LLMs for? Literally the only thing I can think of is generating bad porn of unconsenting people.

Some have mentioned legal work. OpenAI and Anhropic models would refuse to work on cases where something immoral happened.

Re: Grok 4.5

#279
post #5

It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.

The $2/6 pricing seems to only apply for context under 200K. Above that (max context is 500K) pricing doubles to $4/12. https://docs.x.ai/developers/models/grok-4.5

Also, the cache hit pricing is 25% of the input pricing ($2 vs $0.50). Long agentic workflows are dominated by cached input. The US frontier labs typically have this at 10% of the input price, and DeepSeek/Xiaomi etc take it to the extreme 1% range (which is why those are cheap to run in real world agentic loops with dozens of toolcalls per run)

Re: Grok 4.5

#280

Earlier quoted context omitted.

Is uncensored a selling point? What do people use uncensored Grok for (like, real use cases) that they can't or won't use other LLMs for? Literally the only thing I can think of is generating bad porn of unconsenting people.

I don't really have a use for a model that thinks "how many people are in this photo?" is a political question.

I don't know what you're referencing.
Post reply on HN