Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

11–20 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#11

I have never met a single human being who uses Grok for coding

I use. I used to be a Claude user. Since trying Grok 4.5 and especially Grok 4.6, I don't want to go back to Claude any more (I have early access to 4.6).

Grok is 3x+ faster than Claude and I can't tell the diff in engineering work quality. As an engineer, speed is important to me.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#12

I have never met a single human being who uses Grok for coding

I've tried it on my "let's run every model in parallel and see which finds more edge cases" type of tasks, and Grok 4.5 was really behind Opus/ChatGPT but ahead of Gemini - despite having a strong showing on benchmarks.

That makes me really skeptical of it being GPT5.6-tier, much less Fable-tier, based on some of these benchmarks alone. But I'll test here shortly.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#13

I have never met a single human being who uses Grok for coding

I use. I used to be a Claude user. Since trying Grok 4.5 and especially Grok 4.6, I don't want to go back to Claude any more (I have early access to 4.6). Grok is 3x+ faster than Claude and I can't tell the diff in engineering work quality. As an engineer, speed is important to me.

For $30/month, I'd expect it to have higher usage limits than Claude Code and Codex.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#15
post #6

I often wonder if there's a chance, even if minimal... that they stole the weights of the Anthropic models they run on their datacenter... or are actively destillating it.

I think the more likely explanation is that the Cursor data they effectively acquired for $10B was extremely valuable for their training when combined with the insane number of GB300s xAI has for training.

Cursor was 60B. The 10B number was the breakup fee if the deal fell through.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#16

I often wonder if there's a chance, even if minimal... that they stole the weights of the Anthropic models they run on their datacenter... or are actively destillating it.

> ... or are actively destillating it.

I just assumed every model manufacturer is distilling from the frontier models. If they aren't they are definitely trying to do it.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#17

I have never met a single human being who uses Grok for coding

A bunch of SWEs at my work use it as their primary model.

We have Claude, ChatGPT, and Cursor with essentially no cap on spend (top guy is spending over 10K a month on AI at API prices), and he hasn't had his hand slapped.

So it's not like they are using it purely because it's cheaper.

I think people like to use it for its speaking style, pretty solid performance, and its speed.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#19

I have never met a single human being who uses Grok for coding

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business.

Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elon and I don't want to give another dollar to the world's richest person who turns around and uses the money to interfere with elections. The guy I know uses it for essentially the same reason I won't use it.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#20

Earlier quoted context omitted.

I use. I used to be a Claude user. Since trying Grok 4.5 and especially Grok 4.6, I don't want to go back to Claude any more (I have early access to 4.6). Grok is 3x+ faster than Claude and I can't tell the diff in engineering work quality. As an engineer, speed is important to me.

For $30/month, I'd expect it to have higher usage limits than Claude Code and Codex.

It really does. I felt like I could have spent $1000+ api token on claude for the amount of work on my $30 grok subscription.
Post reply on HN