Live data from Hacker News

Grok 4.5

x.ai

181–190 of 1001 posts

Re: Grok 4.5

#181

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

> this doesn’t make sense to me

My hypothesis is that all the top providers realize that, lacking vendor lock in, all SOTA models in a year or so's time will be similar in capability. Also, open weights models are continuing to catch up in a year's time, sometimes less.

So they are trying to lure you in with differentiating, superior capabilities into their proprietary, non-open, non-standard agent harness.

It's the Hotel California playbook: These amazing capabilities are to attract you like moths to a flame and keep you warm and alive around the flame but waterboard and shock you if you attempt to move away from it. Like AWS Egress charges.

Re: Grok 4.5

#182
post #57
post #28

Its remarkable how Anthropic is able to maintain their edge against all competition. Anyone have any idea what the secret sauce is that has Anthropic at the top of all leaderboards for the past few years?

Someone has to know. Would be nice if an insider would drop some hints so that the open-source space could make some good progress.

Nobody has to actually know the secret of their own success, especially not relative success to equally-secretive near-peers.

Same as with rich person autobiographies: even when they tell you what they think it is, they can't see the path not travelled.

Re: Grok 4.5

#183

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

Grok runs tools stupid fast, just about as fast as Antigravity, running Gemini 3.5 Flash.

Re: Grok 4.5

#184

Earlier quoted context omitted.

You could be typing the same about Google or a number of the other labs right now. A diverse market full of choices keeps it from becoming the browser wars all over again.

Google is using AI at such scale internally they don't need external customers to recoup their investment.

> Google is using AI at such scale internally they don't need external customers to recoup their investment.

That's assuming their flagship product remains relevant in an AI-powered world.

Which brings to mind: most of the big shops product (chatgpt, claude, grok, etc...) ALL rely on search, and NONE of them actually have a running search stack.

Which means, they must all be calling Google, no?

How does Google make money from that?

Re: Grok 4.5

#185
post #179

Earlier quoted context omitted.

3rd best chat model? 5th or 6th maybe... GPT Qwen Gemimi MiniMax Claude Ollama GLM Kimi DeepSeek

Ollama is just a local app wrapper/cloud service serving third party apis and models idk why it made it into this list tbh

Claude isn’t a model either.

Re: Grok 4.5

#186
post #8

Of the 3 models I tried, Grok did the best at making an iOS app I wanted for personal use (a bike computer with specific qualities). (Claude just gave up and did an HTML/CSS implementation but I insisted on native SwiftUI+Metal.) Grok definitely fumbles sometimes, but I have been surprised what it CAN intuit versus me having to micromanage it. (I am not an iOS developer, so getting something specific that I needed in…

I swear I have read either this exact or a very similar comment before. Same gist about a bike computer iOS app, and one of the models giving up.

As an aside, big thanks for Caddy! Really helped me get my greenfield project off the ground and it simply “just working” out of the box was one less source of errors I had to worry about when onboarding my team.

Re: Grok 4.5

#187

Earlier quoted context omitted.

Grok is the #1 uncensored easily-available model, and it's also tightly integrated with Twitter.

Is uncensored a selling point? What do people use uncensored Grok for (like, real use cases) that they can't or won't use other LLMs for? Literally the only thing I can think of is generating bad porn of unconsenting people.

I mean absolutely read any thread about Fabel and it's fill with people complaining about how it instantly downgrades or refuses if anything has CVE in the name.

Other then that there is the whole alignment issue. Models that are 'nerfed' in just about any manner tend to exhibit reduced performance is seemingly unrelated areas.

That said Grok doesn't appear to be close enough to the frontier for that to matter. Maybe if they catch up it will.

Re: Grok 4.5

#189
post #3

How popular is Grok compared to other companies models for SWE tasks? I almost never hear it talked about against OpenAI's or Anthropic's products

No one's made a MechaHitler joke yet?

Re: Grok 4.5

#190

Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.

You could be typing the same about Google or a number of the other labs right now. A diverse market full of choices keeps it from becoming the browser wars all over again.

How is this any different than the browser wars? We use to have a diverse market full of choices, and now we have Chromium (almost all market share) and Firefox/Safari on the edges.
Post reply on HN