Its remarkable how Anthropic is able to maintain their edge against all competition. Anyone have any idea what the secret sauce is that has Anthropic at the top of all leaderboards for the past few years?
>Its remarkable how Anthropic is able to maintain their edge against all competition. Anyone have any idea what the secret sauce is that has Anthropic at the top of all leaderboards for the past few years?
It's self-reinforcing: they've got the best coding/research model, which helps them to improve their models better than the competition so they stay ahead.
I have a theory that xAI has one of the largest clusters but with far less traffic + tokens to process bc its less popular than its competition, and xAI can pass the savings on to the end user.
Why would having more costs and less income allow them to pass savings on to the end user?
More like they have a less focus on margins and more on cost recovery.
Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.
[flagged]
I had to scroll down so far to see someone who speaks my language. Thank you. If Grok was the last model on the planet, I would not use it. For the very reason mentioned above. And no, none of the other tech CEOs are that comically evil that they’d take it upon themselves to cut aid from the world’s most vulnerable children while also being the world’s richest man. The optics of that alone… Never letting it go.
So basically since US stopped OpenAI and Anthropic for 4 weeks, it allowed all other AI Labs to almost catch up. GLM 5.2 caught up, Cognition RL'ed Kimi 2.7, Grok 4.5 is out, DeepSeek v4 GA is out in a few days... What is the moat? and why should we pay for the expensive tokens today instead of just waiting a few months/weeks and getting AI for significantly cheaper? I must say, I feel like companies spending Million…
"Almost" is doing a lot of work there; there is no alternative to Fable.
for what it's worth, it's fairly popular among my non-technical coworkers here in Russia. we have unlimited access to all models so it's not about the cost, and they still prefer Gemini over Claude and GPT. I never bothered to ask why, but I assume it's better at communicating in Russian.
This from the country whose entire IT population is still to this day entirely enamored with windows. Not sure it's a valid data point.
to me it seems that IT people overwhelmingly prefer Apple laptops now.
Can someone breakdown to me how this makes any sort of economical sense? Spending billions and billions to have the 3rd best model while even the number 1 and 2 players already seem to struggle making a profit. What am I missing here? Not trying to go full Ed Zitron but this doesn’t make sense to me.
It's simple. Elon's top priority now is "killing the woke mind virus" at any cost, and his Nazi AI is a key tool for that. As long as twitter users take Grok at face value, and spread its talking points all over, it's worth it to him. It doesn't matter if it doesn't make economical sense, it only matters that Elon Musk personally wants to keep it going.
Every time I get excited about Grok’s performance on benchmarks and demo videos, I test it myself and end up disappointed. I'll give this one a try with a grain of salt and lowering my levels of expectations
I'm not - then again I didn't launch a image generation model advertised as having a spicy mode so that might have something to do with the coincidence.
Grok is the #1 uncensored easily-available model, and it's also tightly integrated with Twitter.
Is uncensored a selling point? What do people use uncensored Grok for (like, real use cases) that they can't or won't use other LLMs for? Literally the only thing I can think of is generating bad porn of unconsenting people.
I don't really have a use for a model that thinks "how many people are in this photo?" is a political question.