Live data from Hacker News

Grok 4.5

x.ai

221–230 of 1001 posts

Re: Grok 4.5

#221
post #28

Its remarkable how Anthropic is able to maintain their edge against all competition. Anyone have any idea what the secret sauce is that has Anthropic at the top of all leaderboards for the past few years?

One angle could be their interpretability research? They understand what's going on in LLMs probably much better than anyone else. This must somehow pay off.

I think it's not only an alignment/security tool but could perhaps be used for capabilities as well.

Re: Grok 4.5

#222
post #5

It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.

I have a theory that xAI has one of the largest clusters but with far less traffic + tokens to process bc its less popular than its competition, and xAI can pass the savings on to the end user.

Why would having more costs and less income allow them to pass savings on to the end user?

Re: Grok 4.5

#223

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

Its less about the model; elon is trying to make SpaceXAI a hyper scaler that also happens to have a good model. Grok is just the cherry on top of a powerful AI cluster that can also rent compute to its competitors, like aws.

Re: Grok 4.5

#224
post #5

It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.

I have a theory that xAI has one of the largest clusters but with far less traffic + tokens to process bc its less popular than its competition, and xAI can pass the savings on to the end user.

SpaceX, like Tesla, seems to have the same "portrayals over profits" mindset investors. So it doesn't even really matter whether or not xAI is making any money.

Re: Grok 4.5

#225

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

My guess is that the use here is similar to the reason AWS started as Amazon selling their excess capacity.

Between Tesla, SpaceX, X, Boring Co and Neuralink they probably want the capability internally for a lot of different applications.

If the whole data centers in space thing works out AND people keep protesting/blocking data center build outs on land SpaceX will eventually dominate the entire AI industry just based on escaping scarcity.

Re: Grok 4.5

#226

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

The product is the stock. It is very valuable when you have various bundles of services, such as satellites, AI, and so on, to keep pace with the majors so that you keep pace with their valuation. These stacking valuations are not additive, they're multiplicative because you additionally market investors to the synergy between them. Having the third best model statistically is extremely useful in this context.

The weaknesses can be multiplicative as well. One division bleeding capex can drag down all the rest, no matter how well they might be doing. And the P/E ratio on all of them is riding unrealistic expectations, which can actually be fine for a long time but forces growth even in areas where it doesn't make sense. (Maybe that's where the "let's build data centers in a high radiation hard vacuum!" nonsense comes in; you just need a story of how the P/E ratio is possible to justify in the future? No need to argue over likelihood, just have a tale to tell?)

Re: Grok 4.5

#227
post #144

Earlier quoted context omitted.

You could be typing the same about Google or a number of the other labs right now. A diverse market full of choices keeps it from becoming the browser wars all over again.

Google at least is serving AI results on SRPs billions of times a day, and has pre-existing expertise in data center buildouts and custom silicon. They have one of the more compelling cases for rolling their own.

X has grok built in to every post as does every Tesla Car

Re: Grok 4.5

#228

Earlier quoted context omitted.

Google is using AI at such scale internally they don't need external customers to recoup their investment.

> Google is using AI at such scale internally they don't need external customers to recoup their investment. That's assuming their flagship product remains relevant in an AI-powered world. Which brings to mind: most of the big shops product (chatgpt, claude, grok, etc...) ALL rely on search, and NONE of them actually have a running search stack. Which means, they must all be calling Google, no? How does Google make m…

The funny thing about Google is that Google Search is happy to serve LLM labs search results if it drives their metrics up. Just like Google Cloud is happy to sell off compute to OAI an Anthropic to drive up their metrics.

Google also owns 15% of Anthropic and Hassabis, the leader of Deepmind, also is an early angel investor in Anthropic.

When you really break it down, it's not totally clear that Google would even care that much about being the SOTA LLM.

Re: Grok 4.5

#229
post #222

Earlier quoted context omitted.

I have a theory that xAI has one of the largest clusters but with far less traffic + tokens to process bc its less popular than its competition, and xAI can pass the savings on to the end user.

Why would having more costs and less income allow them to pass savings on to the end user?

They already invested in the massive datacentres of GPUs sitting idle. They have fewer users so they can deliver more inference per user - more thinking, larger models.

Re: Grok 4.5

#230

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

My guess is that the use here is similar to the reason AWS started as Amazon selling their excess capacity. Between Tesla, SpaceX, X, Boring Co and Neuralink they probably want the capability internally for a lot of different applications. If the whole data centers in space thing works out AND people keep protesting/blocking data center build outs on land SpaceX will eventually dominate the entire AI industry just ba…

That Amazon story is a misnomer. They just saw an opportunity with the tech and hardware they had to make a new offering for customers. It's not like they could just offer their spare capacity, then eg at peak US time snatch it back for the retail site
Post reply on HN