Live data from Hacker News

Grok 4.5

x.ai

171–180 of 1001 posts

Re: Grok 4.5

#171

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

People are saying, "There are only a couple of frontier labs. This is a really hard problem and not many people can do it."

Elon's reaction to these kinds of statements is oddly predictable.

Re: Grok 4.5

#172
post #29

Earlier quoted context omitted.

Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.

Every time I use Composer 2.5 I have to spend a bunch of time cleaning up its mistakes. It is unusable compared to GPT 5.4 or 5.5. My time is more valuable that I will use a model that doesn’t f** up my code base.

I think we need to be explicit about the domains we're applying Composer 2.5 to in these discussions.

I mentioned here (https://news.ycombinator.com/item?id=48766275) how poorly it handles my specific use cases. My coworkers in DevOps and frontend UI swear by its cost-effectiveness, whereas I strongly prefer the reasoning capabilities of Opus 4.8 and Fable 5.

Composer 2.5 seems to be SOTA for Helm charts and React/Vue, but, for my usecases it absolutely struggles spectacularly when tasked with rigid body dynamics or kinematic logic.

Re: Grok 4.5

#173

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

All they have to do to differentiate is differentiate the shape of worldview through RLAIF/RHLF and system prompts.

Re: Grok 4.5

#174
What would have been fantastic is if Cursor offered Grok 4.5 in the same usage tier as "Auto + Composer", than provide it as "double usage until July 12" under the API tier (which is what they're doing right now).

EDIT: After looking at my own usage stats - I stand corrected! It is under the "Auto + Composer" tier - brilliant!

Re: Grok 4.5

#175

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

You could be typing the same about Google or a number of the other labs right now. A diverse market full of choices keeps it from becoming the browser wars all over again.

Google is using AI at such scale internally they don't need external customers to recoup their investment.

Re: Grok 4.5

#176

Earlier quoted context omitted.

Grok is the #1 uncensored easily-available model, and it's also tightly integrated with Twitter.

Is uncensored a selling point? What do people use uncensored Grok for (like, real use cases) that they can't or won't use other LLMs for? Literally the only thing I can think of is generating bad porn of unconsenting people.

[deleted]

Re: Grok 4.5

#177
Great model, very nice. Opus class performance at Haiku level pricing (or cheaper with the token efficiency). This seems like a GLM-5.2 killer and this is what Sonnet 5 should have been.

This is a model I could really see used inside applications, where Opus or Sonnet or GPT-5.5 are too expensive.

I would really like to see a strong Deepseek v4-Flash competitor, which ideally is something like Sonnet 4.6 performance at <$0.30 per token. This is missing from main US labs.

Re: Grok 4.5

#178

Earlier quoted context omitted.

[flagged]

What kind of comment is this? Such bad faith and adding nothing to the discussion.

https://www.wired.com/story/grok-is-still-hosting-sexualized...

https://www.pbs.org/newshour/world/musks-grok-chatbot-faces-...

Re: Grok 4.5

#179

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

3rd best chat model? 5th or 6th maybe... GPT Qwen Gemimi MiniMax Claude Ollama GLM Kimi DeepSeek

Ollama is just a local app wrapper/cloud service serving third party apis and models idk why it made it into this list tbh

Re: Grok 4.5

#180
post #121

Earlier quoted context omitted.

Surely grok has a built-in market with too-online, retired boomers. It's free real estate.

This comment says more about your misunderstanding of the world than anything about X

I thought it was pretty accurate tbh.
Post reply on HN