Live data from Hacker News

Grok 4.5

x.ai

41–50 of 1001 posts

Re: Grok 4.5

#41
post #33

Isn't this the same Twitter company that was supposed to go bankrupt a few years ago? Now it is somehow part of a Space company that has an AI division inside of it? I think we are going to be waiting a long time for Twitter / X to go bankrupt as it was (erroneously) predicted a long time ago.

That was the point of the bailout. Twitter is already a rounding error so no one will notice if it goes to zero.

Re: Grok 4.5

#42
post #3

How popular is Grok compared to other companies models for SWE tasks? I almost never hear it talked about against OpenAI's or Anthropic's products

They were missing a harness like Claude Code or Codex (terminal). However they recently released Grok Build, which is probably the fasted I've used, in terms of responsiveness, but didn't have a model at Opus 4.7/8 level. The thing is if they add 4.5 to Grok Build and keep improving the harness I think it can compete (cheaper and faster).

I've been using Grok Build over the last couple weeks. It's actually a very good CLI. The Grok Build 0.1 model isn't great but can also use Composer 2.5 which is excellent. Well worth trying.

Re: Grok 4.5

#43
post #29

Earlier quoted context omitted.

Now if they could have an "equivalent" to Claude's $100 plan with similar compute limits. I have the $40 a month version of Grok and I get a max of like 8 hours of "non-stop" Grok Build coding, per month.

Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.

Composer 2.5 is so underrated IMO. I built a really feature rich application, insanely complicated, close to 200k LOC since it came out and for the most part it ran like a champ. Only used CLaude a couple times to get it unstuck. 8 hours a day and I'm paying about 30 a month.

Re: Grok 4.5

#44
post #28

Its remarkable how Anthropic is able to maintain their edge against all competition. Anyone have any idea what the secret sauce is that has Anthropic at the top of all leaderboards for the past few years?

Given their pricing, I'd guess their models are just way bigger in parameter count. They've always underperformed in cost-per-performance.

They also target a cost-insensitive market (corporate/coding users) compared to Google/OpenAI which support massive amounts of free users.

Re: Grok 4.5

#45

Is there a reason the AI companies usually announce new products so close to each other. Like not just the same day but literally hours apart. GPT Live then an hour later Grok 4.5. As if they try to one up. I expect something new from Anhtropic as well today.

Competition. You don't want to lose your customers trying out the competitors updated and better product. Release on the same day and they won't be able to compare their new to your old.

But how do they know what day is that? Unless you have already something ready to be announced (and you just hold it until the very last moment, which doesn’t make sense, since you could just announce it asap)

Re: Grok 4.5

#46
post #31

Earlier quoted context omitted.

xAI is under criminal investigation in the EU

Who isn't

I'm not - then again I didn't launch a image generation model advertised as having a spicy mode so that might have something to do with the coincidence.

Re: Grok 4.5

#47
post #33

Isn't this the same Twitter company that was supposed to go bankrupt a few years ago? Now it is somehow part of a Space company that has an AI division inside of it? I think we are going to be waiting a long time for Twitter / X to go bankrupt as it was (erroneously) predicted a long time ago.

None of them go bankrupt. The whole thing will just get stuffed into a larger Matryoshka egg that IPOs for eleventy trillion dollars in 10 years.

Re: Grok 4.5

#48
post #8

Of the 3 models I tried, Grok did the best at making an iOS app I wanted for personal use (a bike computer with specific qualities). (Claude just gave up and did an HTML/CSS implementation but I insisted on native SwiftUI+Metal.) Grok definitely fumbles sometimes, but I have been surprised what it CAN intuit versus me having to micromanage it. (I am not an iOS developer, so getting something specific that I needed in…

I do a lot of native iOS development using Opus 4.8 (and I used 4.7/4.6 before this). I have a very hard time with this comment, were you using Opus or something else?

Same. A few months ago I pointed Opus 4.6 at a mid-size Vue app and told it to create the iOS equivalent using SwiftUI, and it nailed it. I broke the process down to phases and reviewed each phase, but within about ten days I had a functioning iOS app that had full feature parity.

Re: Grok 4.5

#49
post #5

It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.

Annoying they didn't show benchmarks for several effort modes, since it seems like it might close the gap with Opus 4.8 by cranking tokens up?

Noam Brown (OpenAI) "Implications of Large-Scale Test-Time Compute" https://xcancel.com/i/article/2064210146558136827

Re: Grok 4.5

#50
post #28

Its remarkable how Anthropic is able to maintain their edge against all competition. Anyone have any idea what the secret sauce is that has Anthropic at the top of all leaderboards for the past few years?

because in the real-world, it's far better than the rest. That's why few people use Grok, it's not even close in day to day work.
Post reply on HN