Earlier quoted context omitted.
Even without the politics, Elon has shown that he will weaponize his platforms against people/companies he personally doesn't like (e.g. specific bans/demotions to external sites like Substack and Bluesky). Using Grok is therefore a supply chain risk and it's not nearly good enough to offset that risk.
I do just want to focus on the 'even without the politics' asterisk though because sometimes there is a risk people think everyone on x side (x meaning 'a given side', not x.com) is wrong You can claim Elon bought x as some sort of power trip. Fine. Willing to entertain it, I have no dog in the fight. I'm not a member of the Elon fan club. And yet Twitter (under Dorsey though I don't think he was involved) was bannin…
Grok 4.5
251–260 of 1001 posts
Re: Grok 4.5
#252Earlier quoted context omitted.
Google is using AI at such scale internally they don't need external customers to recoup their investment.
> Google is using AI at such scale internally they don't need external customers to recoup their investment. That's assuming their flagship product remains relevant in an AI-powered world. Which brings to mind: most of the big shops product (chatgpt, claude, grok, etc...) ALL rely on search, and NONE of them actually have a running search stack. Which means, they must all be calling Google, no? How does Google make m…
Don't they? Based on traffic to some websites I run the big AI labs are very actively doing a lot of crawling.
Re: Grok 4.5
#253Re: Grok 4.5
#254Of the 3 models I tried, Grok did the best at making an iOS app I wanted for personal use (a bike computer with specific qualities). (Claude just gave up and did an HTML/CSS implementation but I insisted on native SwiftUI+Metal.) Grok definitely fumbles sometimes, but I have been surprised what it CAN intuit versus me having to micromanage it. (I am not an iOS developer, so getting something specific that I needed in…
> Claude just gave up and did an HTML/CSS implementation but I insisted on native SwiftUI+Metal. That sounds very odd and very contrary to my experience. You don’t say which model you actually used, but I never had opus 4.8 (or sonnet for that matter) ignore which language/stack i wanted to use.
Re: Grok 4.5
#255Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.
Frontier is one thing, but low-cost really good models are another. All the chatbots and day-to-day corporate bots are likely to use models that offer the best performance at the lowest cost. I think Grok has an angle here if they can build customer trust.
Re: Grok 4.5
#256Do we have any proof that this was made by xAI and isn't some Chinese open model running with modifications? Their inital image generation was a wrapper around Flux.
Re: Grok 4.5
#257Earlier quoted context omitted.
Grok Build sucks compare to composer 2.5. Just use compose 2.5 and you'll have basically unlimited usage on the 40$ plan.
It is hard to evaluate the model performance of Composer 2.5 when Cursor's harness is so awful compared to the others on the market.
Re: Grok 4.5
#258Every time I get excited about Grok’s performance on benchmarks and demo videos, I test it myself and end up disappointed. I'll give this one a try with a grain of salt and lowering my levels of expectations
Re: Grok 4.5
#259It seems to be extremely economical - 4x better reasoning efficiency compared to Opus while being priced at $2/$6. For comparison, GPT 5.4 is $2.5/$15, GPT 5.5/5.6 are $5/$30, Opus 4.8 is $5/$25, Fable is $10/$50. And by benchmarks (unless they gamed them), seems to be at around Opus 4.7 level, which is what Elon mentioned in https://x.com/elonmusk/status/2074911038286295049 . I guess the Cursor data was very useful.
The $2/6 pricing seems to only apply for context under 200K. Above that (max context is 500K) pricing doubles to $4/12. https://docs.x.ai/developers/models/grok-4.5
Re: Grok 4.5
#260Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.
Grok is the #1 uncensored easily-available model, and it's also tightly integrated with Twitter.