Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

101–110 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#101

Cursor, since Grok 4.5, has had an incredible deal for frontier level models, their subscription now goes way further than OpenAI or Anthropic. Even on their lower tier plans you can use a lot tokens on their of their first party models (Grok and Composer) and not really run out comparatively. Combine them with an orchestrator and implementor type setup and it goes even further.

I believe they are the only western provider that has Kimi K3 on a subscription plan today as well. I would love to ditch Anthropic and be on Kimi if there were a subsidized plan like that with ZDR

[deleted]

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#102
post #5

Cursor, since Grok 4.5, has had an incredible deal for frontier level models, their subscription now goes way further than OpenAI or Anthropic. Even on their lower tier plans you can use a lot tokens on their of their first party models (Grok and Composer) and not really run out comparatively. Combine them with an orchestrator and implementor type setup and it goes even further.

Can you explain what you mean? These days courtesy of an addictive reset game OpenAI is playing, I can't find anything with frontier intelligence that's more cost efficient...

If they didn’t constantly reset, they’d be about the same as Anthropic.

Right now, I find that Grok offers better value, uses fewer tokens per turn, and makes better code. I haven’t tried Cursor because I don’t want to change editors again, but maybe I should try it…

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#103
post #87

Earlier quoted context omitted.

Agreed but there were similar controversies with how OpenAI was generating images. The only pass is these are the early days of chatbots and this stuff is so non-deterministic and experimental. For context, this was the change Grok's team made, that was later reverted: > - The response should not shy away from making claims which are politically incorrect, as long as they are well substantiated. https://github.com/xa…

For sure. The difference is that they've made a number of similar suspect changes to Grok on X. Like that weird couple of hours where it would only talk about white genocide in South Africa no matter how you prompted it. Everyone makes mistakes, especially with frontier models. The stuff with Grok shows that the person running the show has a pretty transparent agenda that the company isn't willing to push back on a b…

Yeah I don't care to use Grok's chat and find @grok responses on X mostly noise. I am still open to using it as a backup model for coding though, assuming it does a good job for the price. But mostly because I was already a Cursor user before they bought it.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#105

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

Except EUV lithography is the most complicated industrial process that exists and they won't have usable yields for many years if ever. I don't think Musk actually expects these fans to ever actually make sense they just let him hype and distract.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#106

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

But can you trust any number or metric coming out of SpaceX given everything? Also you mean their own compute like the illegal data centre turbines?

https://www.theguardian.com/technology/2026/jan/15/elon-musk...

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#107

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

> I think they will pull ahead with cheaper tokens similar intelligence

They obviously have a huge token cost advantage of the AI labs they are renting compute to, at least for now while they can charge current crazy rates for GPU compute.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#109
post #19

Earlier quoted context omitted.

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

Versus Sam,Dario or the CCP? Im all for running local models but im sor far from being able to pay for a large model hardware setup. My strix halo box is like driving a beaten up vespa when the frontier models are Ferraris.

[dead]

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#110
post #47

Earlier quoted context omitted.

That's pretty surprising. Idk about Grok 4.6, but Grok 4.5 was clearly below Fable, Opus 5 and GPT 5.6 Sol.

It's much faster, so if you're not doing something cutting-edge, or you're doing the planning yourself and just using the LLM for implementation, the speed benefit outweighs the extra smarts of Fable/Opus5/Sol.

Is there a good way to auto switch between models for planning/implementation?
Post reply on HN