Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

71–80 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#71

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

[flagged]

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#73
post #19

I have never met a single human being who uses Grok for coding

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

[flagged]

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#74

Earlier quoted context omitted.

For $30/month, I'd expect it to have higher usage limits than Claude Code and Codex.

It really does. I felt like I could have spent $1000+ api token on claude for the amount of work on my $30 grok subscription.

I'm really not burning tokens fast enough. I use Claude a lot, daily, and have yet to hit my a ceiling with my Max/100 subscription.

Maybe because I like to verify its outputs and spend a lot of time iterating to get better outcomes. Presumably if I just let it "do its thing" I'd burn more tokens and "get more done" but I'd lose my grasp on what's in the code base.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#75
post #31

Earlier quoted context omitted.

Probably can't advance frontier math yet, yeah. But please let us know other places you want to see Grok improve for future models!

[flagged]

Grok doesn’t have those features, and people who like to make adult content have been complaining for a while how Grok has made it a lot harder to do so.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#76
post #47
post #17

Earlier quoted context omitted.

A bunch of SWEs at my work use it as their primary model. We have Claude, ChatGPT, and Cursor with essentially no cap on spend (top guy is spending over 10K a month on AI at API prices), and he hasn't had his hand slapped. So it's not like they are using it purely because it's cheaper. I think people like to use it for its speaking style, pretty solid performance, and its speed.

That's pretty surprising. Idk about Grok 4.6, but Grok 4.5 was clearly below Fable, Opus 5 and GPT 5.6 Sol.

It's much faster, so if you're not doing something cutting-edge, or you're doing the planning yourself and just using the LLM for implementation, the speed benefit outweighs the extra smarts of Fable/Opus5/Sol.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#77
post #59
post #41

[flagged]

MechaHitler was something that existed only on X's grok chatbot, due to a one-line system prompt change they reverted after half a day. That's different than using Grok as a model for coding.

It's true, but I do worry about governance when it comes to these models. That shows a surprising lack of discipline in their deployment pipeline.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#78
post #31
post #26

Grok is not the best model around, but it's decent. It gets the basic job done at a low price. I don't think it can advance frontier Math, yet.

Probably can't advance frontier math yet, yeah. But please let us know other places you want to see Grok improve for future models!

I recently decided to get an AI subscription and evaluated Grok vs chatgpt. Went with Grok because it's all-around good enough at day to day stuff, integrates with my Tesla, and the image/video generation is great. Kids love whimsical videos of them riding dinosaurs.

Feedback: I'd like Grok to have more connectors (I see OpenAI just added Apple Health, that would be nice to have, and I wish it could read my Onenote notebooks) and for existing ones to be improved. I gave it access to my gmail and asked it "what was my last electricity bill?". It failed to find it, even when I told it the exact subject line to search for. Something about not getting any data back when trying to get the email contents.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#79
post #37

I have never met a single human being who uses Grok for coding

I had a security incident the other day and Grok was the only model that would help. Claude and GPT refused on ethical grounds and only gave general advice. In an emergency, I'd only trust Grok. However, that's the only time I used Grok for coding (since Opus 4.8 it would take a lot to get me to switch away from Anthropic)

With Claude I start having to limit the context I give it about my problem in case it trips up the safeguards. :(

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#80

Cursor, since Grok 4.5, has had an incredible deal for frontier level models, their subscription now goes way further than OpenAI or Anthropic. Even on their lower tier plans you can use a lot tokens on their of their first party models (Grok and Composer) and not really run out comparatively. Combine them with an orchestrator and implementor type setup and it goes even further.

I believe they are the only western provider that has Kimi K3 on a subscription plan today as well. I would love to ditch Anthropic and be on Kimi if there were a subsidized plan like that with ZDR

Opencode have it in their subscription
Post reply on HN