Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

261–270 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#261
I've been using grok 4.5 with grok build soon after it came out and dropped claude. primarily for personal code. It communicates better. While that might not sound like a big deal it is. It doesn't give me a wall of text, tells me what I need to know and I'll make the actual decisions. It is very quick as well which means the sessions are far more interactive, I'll be steering it more. I sometimes cross check with codex and sol, but the daily driver is grok for me.

I found it has improved my productivity and output over claude where it felt like claude was giving me work to do. furthermore with the recent claude watermarking thing, I'd rather use grok or openai.

If anyone is curious download grok cli and throw a couple of prompts at it. you'll be surprised.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#262
post #19

I have never met a single human being who uses Grok for coding

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

everyone interferes with elections. you just prefer people to interfere on your side.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#263
post #113

I'm just wondering why they sold compute to Anthropic if they were planning on still competing in this race?

Competing doesn't mean winning

The rental deal can be terminated by either side with 90 days notice, and presumably Musk would do so if he needed the compute or generally thought it advantageous to do so. For now he doesn't need the compute.

The rental deal may also have been at least in part to juice the SpaceX IPO and to help Anthropic stick it to his enemy OpenAI.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#266
post #140
post #124

Earlier quoted context omitted.

They don't do request based pricing anymore. Its just token based (1 credit = $0.01) plus some bonus credit based on which plan you subscribe. So for example a $39 plan get $70 of credits. https://github.com/features/copilot/plans https://github.blog/changelog/2026-08-06-kimi-k3-is-now-avai...

Yeah, I know, but "credit" translates differently because the models bill at different rates, which gets turned into "multipliers" (or at least, it did). Have they converted entirely to transparent API rates + base allocation now? One of the reasons I left was that if I was going to be billed at API rates anyway , I'd just rather use the APIs. The value proposition still sucks for individuals now, when the other majo…

Yes, they have transparent api rates. And for Anthropic and OpenAI their rates are exactly like API pricing.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#267

Earlier quoted context omitted.

Chinese models are open, Grok is not.

chances are high they are open for now because they infiltrating market and collecting data.

collecting data by making them open weight? where is logic in that?

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#269
post #106

Earlier quoted context omitted.

But can you trust any number or metric coming out of SpaceX given everything? Also you mean their own compute like the illegal data centre turbines ? https://www.theguardian.com/technology/2026/jan/15/elon-musk...

Anthropic rent the same data centre with the turbines from X btw

Wait until they find out what's in a powerplant!

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#270

I've been using grok 4.5 with grok build soon after it came out and dropped claude. primarily for personal code. It communicates better. While that might not sound like a big deal it is. It doesn't give me a wall of text, tells me what I need to know and I'll make the actual decisions. It is very quick as well which means the sessions are far more interactive, I'll be steering it more. I sometimes cross check with co…

Sounds like a caveman skill.
Post reply on HN