Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

341–350 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#341

I've been using grok 4.5 with grok build soon after it came out and dropped claude. primarily for personal code. It communicates better. While that might not sound like a big deal it is. It doesn't give me a wall of text, tells me what I need to know and I'll make the actual decisions. It is very quick as well which means the sessions are far more interactive, I'll be steering it more. I sometimes cross check with co…

Is grok build the same as grok cli?

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#342

Cursor, since Grok 4.5, has had an incredible deal for frontier level models, their subscription now goes way further than OpenAI or Anthropic. Even on their lower tier plans you can use a lot tokens on their of their first party models (Grok and Composer) and not really run out comparatively. Combine them with an orchestrator and implementor type setup and it goes even further.

The value in their subscription is bound to the Cursor agent/software only though correct?

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#343

I've been using grok 4.5 with grok build soon after it came out and dropped claude. primarily for personal code. It communicates better. While that might not sound like a big deal it is. It doesn't give me a wall of text, tells me what I need to know and I'll make the actual decisions. It is very quick as well which means the sessions are far more interactive, I'll be steering it more. I sometimes cross check with co…

Out of interest, do Musk's politics impact your decision on whether or not to use Grok? I'd be interested to know where folks lie on the (Agree / Disagree) and (Use / Don't use) axes.

Interesting question. I suppose it comes down to how much you allocate his involvement or presence to a product? I’d imagine Grok is built by hundreds of engineers who are all unique individuals from various backgrounds. If Elon simply “leads” from a very surface level where he has no direct day to day involvement in Grok releases does that make it more palatable? Or is the question really about how involved he is? Or is simply being the leader (even if he was 100% absent and only had his name attached to a project/company) enough to boycott?

On a similar note, how much Elon hate is about his politics vs his trillionaire status vs what I like to call “watercooler hate” where folks simply parrot the loudest opinion in order to be accepted into the group?

On a final note, my son is in primary school and recently brought up in a dinner time discussion that “Elon is really bad” - this is a kid who has no social media (unlike some of his peers who are already on TikTok) and doesn’t watch traditional media.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#344
post #250

Earlier quoted context omitted.

But still for US frontier you're paying 10-20x more per token compared to their limited subscriptions. For China frontier you'll be good though, and that might be the future anyway.

Grok is cheaper vs real Chinese frontier aka kimi. Sponsored or not.

What's the cheapest way to use grok models for coding?

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#345

Earlier quoted context omitted.

In some circles any AI usage puts you in the same bucket as the worst of the worst

In some circles using soap used to put you in the worst of the worst

You’re right; that’s why REST is now popular.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#346

Earlier quoted context omitted.

> many of us don’t want to be repeatedly beaten over the head with other people’s politics I agree completely. That's why I won't use Grok: its owner repeatedly beats us over the head with his politics, and I won't encourage it.

Huh. I feel that way with the other models - all the same. Eh, what can you do. Have fun out there.

Do you really think OpenAI's leadership are as stridently, overtly political as Musk? OK.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#347
Grok is quite interesting. I run comparisons almost daily on tasks and Grok is its own beast, in a good way.

It's good to have model diversity. When I run a task across Sol, Terra, and Luna, I get variations of the same thing with diminishing quality. It makes the lineup pointless. Ditto for Anthropic. Gemini-3.6-Flash and 3.1 Pro genuinely behave differently. Opus 5 and Fable are.. cousins.

I find that when I want to test a complex creative challenge, having 4 "families" to choose from makes the experience interesting since they will excel in different areas.

Grok might implement unique lighting, Opus, elegant primitives, Sol, accurate snowfall in one pass, Gemini, silky movement. Combined, you can pick and choose best.

For what its worth, Grok always feels "messy" but finishes. Grok 4.6 though is no longer "smart and fast". It's about as fast as Sol though.

A big improvement I noticed in 4.6 was tool use for verification. Previously, Opus/Fable were the only models to consistently screenshot things that they can't directly interact with easily. Now Grok is probably right behind them, perhaps tied with Sol on propensity to verify visually. Grok 4.5 notably did not do this often.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#348
post #125

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

OpenAI is almost there, and Anthropic is pretty close behind. In the next year or two all major AI companies will be vertically integrated to a good degree.

Neither of them are "almost there"...

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#349

Earlier quoted context omitted.

"everythin you don't like" is not a scientific argument, neither is the illegality of data centers on AI quality

So you’re asserting that the ends justify the means or? I’m confused by how your reply makes sense in the context of the parent comment. They weren’t stating a preference, they were linking to simple facts. Please do better.

> So you’re asserting that the ends justify the means or?

They claim that spacex has competitive advantage. They take no stance on condoning spacex’s behavior. Personally I strongly dislike musks behavior, but I appreciate discussion of competitive advantages/disadvantages, independent of moral views. I liked the thread, it adds new info to the conversation.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#350
post #119

Reading the SWE bickering back and fourth in this thread about Claude vs Grok reminds me of IE vs Netscape bickering way back when.

Netscape 4 Lyf

It still survives in the cookie jar format.
Post reply on HN