Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

331–340 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#331
post #187

Earlier quoted context omitted.

>> I think they will pull ahead with cheaper tokens similar intelligence they just increased cache read from 0.30 to 0.50 - this has the biggest impact on agentic coding. Elon companies have the most expensive everything: xAI sub: $30 when other starts at $20, pro like sub for $300 where other charge $200. Expensive electric cars, powerwalls, solar roofs when competetive products/better are cheaper.

>Elon companies have the most expensive everything The Model 3 and Model Y became the highest selling EVs of all time because they were the first below $50K to have long-range and be worth buying. Until a few years ago, every other sub-$50K EV absolutely sucked.

> The Model 3 and Model Y became the highest selling EVs of all time because they were the first below $50K to have long-range and be worth buying.

And then Musk totally abandoned Tesla's original brilliant game plan of using the luxury models to find actual low-cost models (which $50K is not), and completely ceded the future EV market to Chinese companies that understand how to make a better car than Tesla for less money. BYD sells more cars than Tesla.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#332

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

> and soon chip making factory

To say that I somewhat doubt this would be an understatement.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#333

Earlier quoted context omitted.

Out of interest, do Musk's politics impact your decision on whether or not to use Grok? I'd be interested to know where folks lie on the (Agree / Disagree) and (Use / Don't use) axes.

[flagged]

In some circles any AI usage puts you in the same bucket as the worst of the worst

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#334

I've been using grok 4.5 with grok build soon after it came out and dropped claude. primarily for personal code. It communicates better. While that might not sound like a big deal it is. It doesn't give me a wall of text, tells me what I need to know and I'll make the actual decisions. It is very quick as well which means the sessions are far more interactive, I'll be steering it more. I sometimes cross check with co…

Out of interest, do Musk's politics impact your decision on whether or not to use Grok? I'd be interested to know where folks lie on the (Agree / Disagree) and (Use / Don't use) axes.

I disagree with Musk's politics but it does not impact my decision to use Grok. That's because being serious about aligning my capital to my values doesn't leave much in the way of eligible products or services. I consequently decide not to worry about this as a moral axis for my life.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#335

Earlier quoted context omitted.

I'm really not burning tokens fast enough. I use Claude a lot, daily, and have yet to hit my a ceiling with my Max/100 subscription. Maybe because I like to verify its outputs and spend a lot of time iterating to get better outcomes. Presumably if I just let it "do its thing" I'd burn more tokens and "get more done" but I'd lose my grasp on what's in the code base.

Big same here with Claude but I’m on the $200 sub. I let Fable run for roughly 4 hours and still didn’t hit the session limit. I suppose if I was running more in parallel it would be easier to hit, but I’m not particularly good at focusing on more than one thing at the same time (even if I’m waiting on an agent to do the work).

I hear you. I usually have two things going at once (either two different machines, two different projects, or different work trees if the same repo), but more than that I don't feel like I'm locked in enough to providing the "thinking" that Claude definitely still needs.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#336

Earlier quoted context omitted.

I've found Fable 5 to be so much better than 4.8. For building a full stack custom CRM and media pipeline tool with video conversion, transcription, and indexing. Supabase, AWS, Meili, NextJS, GCS - lots of surfaces and planes. 4.8 basically couldn't do it, I abandoned the project as the fallback was, "current business processes". With F5 it's been 4 weeks and almost ready for production release.

I have the same quality results with Fable. With just a brief prompt, it created a great static website with a beautiful animation of a workflow. Gemini's output was so poor that I closed the chat. And with Codex, the results were bad, so I discarded them.

Agree. Opus 4.8 could make nice little toy and demo apps. This app had a lot of surfaces and pipeline, Postgres, vector search, S3 -- couldn't handle that. Fable 5 is still a lot of work and you have to check it, but it really does perform at senior eng level. Shipping good size features daily.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#337
post #19

Earlier quoted context omitted.

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

everyone interferes with elections. you just prefer people to interfere on your side.

Every accusation is a confession...

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#338

I've been using grok 4.5 with grok build soon after it came out and dropped claude. primarily for personal code. It communicates better. While that might not sound like a big deal it is. It doesn't give me a wall of text, tells me what I need to know and I'll make the actual decisions. It is very quick as well which means the sessions are far more interactive, I'll be steering it more. I sometimes cross check with co…

I've never used Grok, but I'm very dissatisfied with the writing style of frontier models from OpenAI and Anthropic. I only use them for coding now.

ChatGPT is very long-winded, sometimes producing multiple bullet point lists for a simple answer. Claude is full of mannerisms: 'not merely x, but y', 'Here's where it gets interesting', 'the real question is', etc.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#339

Earlier quoted context omitted.

[flagged]

In some circles any AI usage puts you in the same bucket as the worst of the worst

In some circles using soap used to put you in the worst of the worst

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#340

Earlier quoted context omitted.

Out of interest, do Musk's politics impact your decision on whether or not to use Grok? I'd be interested to know where folks lie on the (Agree / Disagree) and (Use / Don't use) axes.

I avoid Grok for meaningful token spend on purpose/boycotting. I do check in via openrouter occasionally to check it's chat performance which has seemed fine to me since 4. My total grok spend has been ~$2. I disagree with his politics to a huge degree. My token spend at api rates is about $3000 usd a month recently.

Which part of his politics?
Post reply on HN