Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

121–130 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#121
post #94

Earlier quoted context omitted.

> many of us don’t want to be repeatedly beaten over the head with other people’s politics I agree completely. That's why I won't use Grok: its owner repeatedly beats us over the head with his politics, and I won't encourage it.

It's fine to make your own choices about what you want to spend your money on. But it gets old to constantly have actual technical discussion drowned out by the same repetitive low-effort comments in every single post. I want to hear about what people have used Grok for, how it compares to other models, what it's not good at, etc., without having to wade through all the predictable "Elon bad" comments. @dang - featur…

There are already features that let you skip conversations you don't like. Just click the - next to the post and stop whining about it

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#123
post #59

Earlier quoted context omitted.

MechaHitler was something that existed only on X's grok chatbot, due to a one-line system prompt change they reverted after half a day. That's different than using Grok as a model for coding.

[flagged]

I think "system prompt" is the key bit they're getting at. It doesn't necessarily reflect poorly on the underlying model if the system prompt was bad. It does reflect somewhat, in terms of alignment (how well the model does what the training company wants) and instruction following (how well the model does what the user wants). But it's not so clear to me what exactly the right answer is here. E.g., a model that scrupulously follows its system prompt and does what the user wants is a pretty useful, if very sharp, tool, albeit perhaps dangerous in the wrong hands.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#124
post #92
post #44

Earlier quoted context omitted.

GitHub Copilot does have Kimi K3.

What’s the multiplier? GH copilot nerfed their product so badly that I unsubscribed.

They don't do request based pricing anymore. Its just token based (1 credit = $0.01) plus some bonus credit based on which plan you subscribe. So for example a $39 plan get $70 of credits.

https://github.com/features/copilot/plans

https://github.blog/changelog/2026-08-06-kimi-k3-is-now-avai...

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#125

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

OpenAI is almost there, and Anthropic is pretty close behind. In the next year or two all major AI companies will be vertically integrated to a good degree.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#126
post #106

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

But can you trust any number or metric coming out of SpaceX given everything? Also you mean their own compute like the illegal data centre turbines ? https://www.theguardian.com/technology/2026/jan/15/elon-musk...

being a public company forces a lot of trust and transparency bc otherwise shareholders will sue you into oblivion

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#127
post #102
post #5

Earlier quoted context omitted.

Can you explain what you mean? These days courtesy of an addictive reset game OpenAI is playing, I can't find anything with frontier intelligence that's more cost efficient...

If they didn’t constantly reset, they’d be about the same as Anthropic. Right now, I find that Grok offers better value, uses fewer tokens per turn, and makes better code. I haven’t tried Cursor because I don’t want to change editors again, but maybe I should try it…

Are there any projects that track how much usage of each model translates to how much percentage drop in weekly/5hr windows?

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#128

I have never met a single human being who uses Grok for coding

Back in the day (in AI time) GitHub Copilot had Grok on the 0 github-token cost and I found it to be the best of the 0 github-token models for when my budget was out. Then they went to a multiplier that was not competitive and I haven't look back again. Been meaning too, but for personal use, Deepseek flash is so cheap I haven't felt like spending money elsewhere.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#129
post #89

Earlier quoted context omitted.

[flagged]

[flagged]

Idk I just don't like the richest man on planet and the owner of the "town square of the internet" to post fake news blatantly to promote hate against a group of people.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#130
post #59

Earlier quoted context omitted.

MechaHitler was something that existed only on X's grok chatbot, due to a one-line system prompt change they reverted after half a day. That's different than using Grok as a model for coding.

It's true, but I do worry about governance when it comes to these models. That shows a surprising lack of discipline in their deployment pipeline.

yes exactly. if a company is happy to have their LLM's produce neo nazi content and CSAM, why do I want to give them money and my most important digital material?
Post reply on HN