Live data from Hacker News

Grok 4.5

x.ai

251–260 of 1001 posts

Re: Grok 4.5

#251

Earlier quoted context omitted.

Even without the politics, Elon has shown that he will weaponize his platforms against people/companies he personally doesn't like (e.g. specific bans/demotions to external sites like Substack and Bluesky). Using Grok is therefore a supply chain risk and it's not nearly good enough to offset that risk.

I do just want to focus on the 'even without the politics' asterisk though because sometimes there is a risk people think everyone on x side (x meaning 'a given side', not x.com) is wrong You can claim Elon bought x as some sort of power trip. Fine. Willing to entertain it, I have no dog in the fight. I'm not a member of the Elon fan club. And yet Twitter (under Dorsey though I don't think he was involved) was bannin…

Pre-Musk twitter isn't the comparison point here. Anthropic/Google are.

Re: Grok 4.5

#252

Earlier quoted context omitted.

Google is using AI at such scale internally they don't need external customers to recoup their investment.

> Google is using AI at such scale internally they don't need external customers to recoup their investment. That's assuming their flagship product remains relevant in an AI-powered world. Which brings to mind: most of the big shops product (chatgpt, claude, grok, etc...) ALL rely on search, and NONE of them actually have a running search stack. Which means, they must all be calling Google, no? How does Google make m…

> Which brings to mind: most of the big shops product (chatgpt, claude, grok, etc...) ALL rely on search, and NONE of them actually have a running search stack.

Don't they? Based on traffic to some websites I run the big AI labs are very actively doing a lot of crawling.

Re: Grok 4.5

#254
post #8

Of the 3 models I tried, Grok did the best at making an iOS app I wanted for personal use (a bike computer with specific qualities). (Claude just gave up and did an HTML/CSS implementation but I insisted on native SwiftUI+Metal.) Grok definitely fumbles sometimes, but I have been surprised what it CAN intuit versus me having to micromanage it. (I am not an iOS developer, so getting something specific that I needed in…

> Claude just gave up and did an HTML/CSS implementation but I insisted on native SwiftUI+Metal. That sounds very odd and very contrary to my experience. You don’t say which model you actually used, but I never had opus 4.8 (or sonnet for that matter) ignore which language/stack i wanted to use.

It never happened to me, but Claude routinely ignores the single line I have in CLAUDE.md, so I wouldn't be entirely surprised.

Re: Grok 4.5

#255
post #162

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

Frontier is one thing, but low-cost really good models are another. All the chatbots and day-to-day corporate bots are likely to use models that offer the best performance at the lowest cost. I think Grok has an angle here if they can build customer trust.

Quitting my job if I have to use any Musk product… I know Anthropic’s lease of xAI data centers pumped SpaceX stock, so they’re kinda in-bed with each other, but directly using Musk products is pure immorality IMO. Using a Nazi’s products is not an acceptable outcome to me, and I’m fully prepared to change my job/entire career over it. I’m still young, and have time to pivot.

Re: Grok 4.5

#256

Do we have any proof that this was made by xAI and isn't some Chinese open model running with modifications? Their inital image generation was a wrapper around Flux.

It’d be real funny if this was just GLM 5.2 trained on Cursor data

Re: Grok 4.5

#257
post #29

Earlier quoted context omitted.

Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.

It is hard to evaluate the model performance of Composer 2.5 when Cursor's harness is so awful compared to the others on the market.

Not true. The only issue is cost of frontier models.

Re: Grok 4.5

#258

Every time I get excited about Grok’s performance on benchmarks and demo videos, I test it myself and end up disappointed. I'll give this one a try with a grain of salt and lowering my levels of expectations

[flagged]

Re: Grok 4.5

#259
post #5

It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.

The $2/6 pricing seems to only apply for context under 200K. Above that (max context is 500K) pricing doubles to $4/12. https://docs.x.ai/developers/models/grok-4.5

That's very notable and left out of the announcement.

Re: Grok 4.5

#260

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

Grok is the #1 uncensored easily-available model, and it's also tightly integrated with Twitter.

I don't remember online discourses on filter avoidance for Grok to be any different from typical ones, except that it allegedly have tendency to take porn-biased interpretations of prompts, I think the "uncensored" pitch they had for a while was pure marketing in the end.
Post reply on HN