Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

221–230 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#221

Earlier quoted context omitted.

I would be curious to see the venn diagram with the people who think data centers are using all of the water.

A perfect circle. You can probably guess about 99% of the views of the people in that circle. They require strict conformity to the religion.

I've heard it said that "problematic" is the "blasphemous" of said religion.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#222
post #42

Earlier quoted context omitted.

It's a bit worse. I haven't tried so it's pure speculation based on benchmarks, but I'd assume Grok 4.6 is around Opus 4.8 in real world use, but clearly below Opus 5.

I've found Fable 5 to be so much better than 4.8. For building a full stack custom CRM and media pipeline tool with video conversion, transcription, and indexing. Supabase, AWS, Meili, NextJS, GCS - lots of surfaces and planes. 4.8 basically couldn't do it, I abandoned the project as the fallback was, "current business processes". With F5 it's been 4 weeks and almost ready for production release.

I have the same quality results with Fable. With just a brief prompt, it created a great static website with a beautiful animation of a workflow. Gemini's output was so poor that I closed the chat. And with Codex, the results were bad, so I discarded them.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#223

Cursor, since Grok 4.5, has had an incredible deal for frontier level models, their subscription now goes way further than OpenAI or Anthropic. Even on their lower tier plans you can use a lot tokens on their of their first party models (Grok and Composer) and not really run out comparatively. Combine them with an orchestrator and implementor type setup and it goes even further.

>their subscription now goes way further than OpenAI or Anthropic.

Until it doesn't...

Honestly, this entire OpenAI reset credit fiasco this past week has convinced me to rip off the Codex and Claude Code bandaids and start building my own proper Pi Coding Agent running models that I select and pay for on openrouter.

And I am feeling a lot better about it now that I've finally got it working.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#225
post #163
post #102

Earlier quoted context omitted.

If they didn’t constantly reset, they’d be about the same as Anthropic. Right now, I find that Grok offers better value, uses fewer tokens per turn, and makes better code. I haven’t tried Cursor because I don’t want to change editors again, but maybe I should try it…

That is not true; GPT is the most reasoning efficient model family on the market.

Yeah, even without the resets, chatgpt subscription currently goes quite a bit further than an equivalent anthropic plan. The main reason to have an anthropic plan is to get access to Fable 5 if you feel the quality of output makes it worth it.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#226
post #31
post #26

Grok is not the best model around, but it's decent. It gets the basic job done at a low price. I don't think it can advance frontier Math, yet.

Probably can't advance frontier math yet, yeah. But please let us know other places you want to see Grok improve for future models!

Not the model but the app - with Claude I can work on my mac using Cloud environments, leave the office, open Claude on my phone and respond. I can't change model if I have to stop that workflow.

Does xAI have plans to do Mac/iOS apps with cloud environments? When can we expect them?

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#227
post #39

Earlier quoted context omitted.

I can respect if you say you hate their guts. Everyone has their worldview. But over moral or ethical stand? You don't have any if you're using Chinese models, or fly Middle East airlines, or countless of other products. Don't delude yourself.

Chinese models are open, Grok is not.

chances are high they are open for now because they infiltrating market and collecting data.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#228
post #106

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

But can you trust any number or metric coming out of SpaceX given everything? Also you mean their own compute like the illegal data centre turbines ? https://www.theguardian.com/technology/2026/jan/15/elon-musk...

Anthropic rent the same data centre with the turbines from X btw

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#229
post #78

Earlier quoted context omitted.

I recently decided to get an AI subscription and evaluated Grok vs chatgpt. Went with Grok because it's all-around good enough at day to day stuff, integrates with my Tesla, and the image/video generation is great. Kids love whimsical videos of them riding dinosaurs. Feedback: I'd like Grok to have more connectors (I see OpenAI just added Apple Health, that would be nice to have, and I wish it could read my Onenote n…

None of that has anything to do with the model - your issues are all with the harness/app you used...

Yes. Grok is a harness/app as well as a model.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#230

Cursor, since Grok 4.5, has had an incredible deal for frontier level models, their subscription now goes way further than OpenAI or Anthropic. Even on their lower tier plans you can use a lot tokens on their of their first party models (Grok and Composer) and not really run out comparatively. Combine them with an orchestrator and implementor type setup and it goes even further.

>their subscription now goes way further than OpenAI or Anthropic. Until it doesn't... Honestly, this entire OpenAI reset credit fiasco this past week has convinced me to rip off the Codex and Claude Code bandaids and start building my own proper Pi Coding Agent running models that I select and pay for on openrouter. And I am feeling a lot better about it now that I've finally got it working.

But still for US frontier you're paying 10-20x more per token compared to their limited subscriptions. For China frontier you'll be good though, and that might be the future anyway.
Post reply on HN