Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

81–90 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#81
post #42
post #34

Earlier quoted context omitted.

How does Grok 4.5 compare to Opus >= 4.8 though? I'm willing to pay 2x for a 10% smarter model. Intelligence matters that much (because 10% smarter probably saves, on average, several hours of human time).

It's a bit worse. I haven't tried so it's pure speculation based on benchmarks, but I'd assume Grok 4.6 is around Opus 4.8 in real world use, but clearly below Opus 5.

I've found Fable 5 to be so much better than 4.8.

For building a full stack custom CRM and media pipeline tool with video conversion, transcription, and indexing. Supabase, AWS, Meili, NextJS, GCS - lots of surfaces and planes.

4.8 basically couldn't do it, I abandoned the project as the fallback was, "current business processes".

With F5 it's been 4 weeks and almost ready for production release.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#82

I have never met a single human being who uses Grok for coding

I've tried it on my "let's run every model in parallel and see which finds more edge cases" type of tasks, and Grok 4.5 was really behind Opus/ChatGPT but ahead of Gemini - despite having a strong showing on benchmarks. That makes me really skeptical of it being GPT5.6-tier, much less Fable-tier, based on some of these benchmarks alone. But I'll test here shortly.

It's still not as good as GPT5.6 or Opus5 but it's better than KimiK3. Good job xAI team.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#83
post #19

I have never met a single human being who uses Grok for coding

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

Versus Sam,Dario or the CCP? Im all for running local models but im sor far from being able to pay for a large model hardware setup. My strix halo box is like driving a beaten up vespa when the frontier models are Ferraris.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#84
post #19

Earlier quoted context omitted.

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

[flagged]

> MADISON, Wis. (AP) — Billionaire Elon Musk likely broke Wisconsin law when he promised to hand out $1 million checks to voters in the 2025 state Supreme Court election, a bipartisan panel has found.

> The Wisconsin Elections Commission last week referred two complaints to the Brown County district attorney’s office, which can choose to bring criminal charges over violating the state law against election bribery. Prosecutors have 40 days to report back to the commission.

https://apnews.com/article/elon-musk-wisconsin-election-mill...

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#85
post #19

Earlier quoted context omitted.

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

[flagged]

This statement doesn't go far enough given Elon's direct and hands-on involvement with DOGE and the 2024 elections. Very few of the richest people of the world are personally entangled in meddling with government agencies directly, for example.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#86

Earlier quoted context omitted.

[flagged]

Grok doesn’t have those features, and people who like to make adult content have been complaining for a while how Grok has made it a lot harder to do so.

[flagged]

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#87
post #59

Earlier quoted context omitted.

MechaHitler was something that existed only on X's grok chatbot, due to a one-line system prompt change they reverted after half a day. That's different than using Grok as a model for coding.

It's true, but I do worry about governance when it comes to these models. That shows a surprising lack of discipline in their deployment pipeline.

Agreed but there were similar controversies with how OpenAI was generating images. The only pass is these are the early days of chatbots and this stuff is so non-deterministic and experimental.

For context, this was the change Grok's team made, that was later reverted:

> - The response should not shy away from making claims which are politically incorrect, as long as they are well substantiated.

https://github.com/xai-org/grok-prompts/commit/c5de4a14feb50...

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#88
post #19

Earlier quoted context omitted.

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

>who turns around and uses the money to interfere with elections thats a very dumb reason considering all rich people do it, most are just not as open about it as Musk

There's a big difference between quiet donations to a PAC - not that that's good either - and what Elon did. He literally paid for votes, likely in violation of the law. He poured more money into U.S. elections than anyone has before. He and Trump both made strange, cryptic statements about Elon's role in Pennsylvania with the voting machines that has caused people to reasonably wonder if they somehow manipulated the election. Whether he did or not, the innuendo alone is not ok. Then he did what he did in Germany. Don't even get me started on DOGE or his Starlink shenanigans in Ukraine.

That's before we even start talking about the models themselves. He claims to want "unbiased" models, but he very clearly has a distorted view of the world and has repeatedly demonstrated a desire and willingness to bend the world to his will. I don't want to use a model that is so obviously suspect. Not to mention, his models repeatedly produce racist, Nazi-like propaganda.

IMO, he is, at best, a clueless amateur masquerading as an expert and running into problems a more careful person manages to mostly avoid. At worst... well, you get the picture.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#89
post #19

Earlier quoted context omitted.

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

[flagged]

[flagged]
Post reply on HN