Live data from Hacker News

Grok 4.5

x.ai

61–70 of 1001 posts

Re: Grok 4.5

#61
Every time I get excited about Grok’s performance on benchmarks and demo videos, I test it myself and end up disappointed.

I'll give this one a try with a grain of salt and lowering my levels of expectations

Re: Grok 4.5

#62
post #29

Earlier quoted context omitted.

Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.

Composer 2.5 is so underrated IMO. I built a really feature rich application, insanely complicated, close to 200k LOC since it came out and for the most part it ran like a champ. Only used CLaude a couple times to get it unstuck. 8 hours a day and I'm paying about 30 a month.

> I built a really feature rich application, insanely complicated, close to 200k LOC

If you listed it, how many features/LOC or vice-versa? Really hard to know if 200K LOC is good or bad, at the surface it sounds like too much, but I don't know what the application was either.

Re: Grok 4.5

#63
post #28

Its remarkable how Anthropic is able to maintain their edge against all competition. Anyone have any idea what the secret sauce is that has Anthropic at the top of all leaderboards for the past few years?

My gut feel is Anthropic is very technical and pedantic which makes their models really technical and pedantic. They're top at code and technical benchmarks but anecdotally I've found OpenAI to be significantly farther ahead for general usage.

Opus 4.8 will burn 10k tokens trying to answer something 100% whereas GPT-5.5 will burn 2k getting it 90% which is good enough for many things.

Some personal testing on a "help me find that restaurant" prompt https://gist.github.com/nijave/2873b8b10d8c732e46264237b0755...

Re: Grok 4.5

#65

Earlier quoted context omitted.

Around Opus 4.7 level would be the same as Sonnet 5 while being cheaper overall. I wonder how good their subscription discount is on both their subscription types.

Sonnet 5 is a huge token hog, though, it uses far more reasoning tokens than Opus models while being priced at $2/$10 with promo, and $3/$15 (usual Sonnet price) afterwards.

I'll probably get hate for it, but I was not impressed by Fable, I felt like it was just Opus with more tokens for thinking. I feel like the second I turned on Fable I drained my usage more quickly, despite them billing it as though it were Opus level of usage. The value is just not there for me. I wish they could make Haiku remain low-cost and drastically more capable to the point you could use only Haiku.

Re: Grok 4.5

#66
post #55

[flagged]

Low effort and uninformed comment. The team published a good followup on why this happened: the model pulled in people's own tweets as context to prompting so edge lords that wrote innocuous prompts got to see edge lord content.

Re: Grok 4.5

#67
post #29

Earlier quoted context omitted.

Now if they could have an "equivalent" to Claude's $100 plan with similar compute limits. I have the $40 a month version of Grok and I get a max of like 8 hours of "non-stop" Grok Build coding, per month.

Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.

Every time I use Composer 2.5 I have to spend a bunch of time cleaning up its mistakes. It is unusable compared to GPT 5.4 or 5.5.

My time is more valuable that I will use a model that doesn’t f** up my code base.

Re: Grok 4.5

#68

Is there a reason the AI companies usually announce new products so close to each other. Like not just the same day but literally hours apart. GPT Live then an hour later Grok 4.5. As if they try to one up. I expect something new from Anhtropic as well today.

Maybe it‘s the Nash equilibrium from a timing perspective?

Like the reason that close to a McDonals there is usually a Burger King.

Re: Grok 4.5

#69
post #28

Its remarkable how Anthropic is able to maintain their edge against all competition. Anyone have any idea what the secret sauce is that has Anthropic at the top of all leaderboards for the past few years?

I think it's the talent, laser focus on single product set and being early so ahead, same with Open AI who are only a sliver behind. Google, XAI are the next level down but they have other concerns.

Re: Grok 4.5

#70
So basically since US stopped OpenAI and Anthropic for 4 weeks, it allowed all other AI Labs to almost catch up.

GLM 5.2 caught up, Cognition RL'ed Kimi 2.7, Grok 4.5 is out, DeepSeek v4 GA is out in a few days...

What is the moat? and why should we pay for the expensive tokens today instead of just waiting a few months/weeks and getting AI for significantly cheaper?

I must say, I feel like companies spending Millions on Anthropic tokens are just negative capex'ing and wasting money, even OpenAI is barely ok pricing...

Post reply on HN