How popular is Grok compared to other companies models for SWE tasks? I almost never hear it talked about against OpenAI's or Anthropic's products
Grok 4.5
211–220 of 1001 posts
Re: Grok 4.5
#212Earlier quoted context omitted.
Sonnet 5 is a huge token hog, though, it uses far more reasoning tokens than Opus models while being priced at $2/$10 with promo, and $3/$15 (usual Sonnet price) afterwards.
I'll probably get hate for it, but I was not impressed by Fable, I felt like it was just Opus with more tokens for thinking. I feel like the second I turned on Fable I drained my usage more quickly, despite them billing it as though it were Opus level of usage. The value is just not there for me. I wish they could make Haiku remain low-cost and drastically more capable to the point you could use only Haiku.
Simple tasks are simply saturated just like simple benchmarks. There's a level of intelligence where you simply don't need more for some things.
Re: Grok 4.5
#213Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.
SpaceX needs to keep raising many billions every year. The rockets part isn't going to make money for a long time, so diversion tactics https://news.ycombinator.com/item?id=48828648 Also Elon has a grudge with Sam Altman and wants to beat him
Re: Grok 4.5
#214Very hard for me to imagine this getting beyond a low-single-digit market share. I don't understand the strategy of xAI burning money on this.
Re: Grok 4.5
#215It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.
Re: Grok 4.5
#216Do we have any proof that this was made by xAI and isn't some Chinese open model running with modifications? Their inital image generation was a wrapper around Flux.
Genuinely asking.
Re: Grok 4.5
#217Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.
Grok is the #1 uncensored easily-available model, and it's also tightly integrated with Twitter.
Re: Grok 4.5
#218Re: Grok 4.5
#219Re: Grok 4.5
#220Earlier quoted context omitted.
Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.
It is hard to evaluate the model performance of Composer 2.5 when Cursor's harness is so awful compared to the others on the market.