Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

231–240 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#231
post #10

I have never met a single human being who uses Grok for coding

Folks working in US govt tend to, based on convos I've had with one such person.

Grok probably doesn’t object when the government asks it how to bomb schools.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#233
post #187

SpaceXAI is the only frontier model company that had its own compute/date centres and soon chip making factory, I think they will pull ahead with cheaper tokens similar intelligence and better harness/tools. Grok build is 2-5x faster than Claude Code in my opinion.

>> I think they will pull ahead with cheaper tokens similar intelligence they just increased cache read from 0.30 to 0.50 - this has the biggest impact on agentic coding. Elon companies have the most expensive everything: xAI sub: $30 when other starts at $20, pro like sub for $300 where other charge $200. Expensive electric cars, powerwalls, solar roofs when competetive products/better are cheaper.

>Elon companies have the most expensive everything

The Model 3 and Model Y became the highest selling EVs of all time because they were the first below $50K to have long-range and be worth buying.

Until a few years ago, every other sub-$50K EV absolutely sucked.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#234
post #152

Earlier quoted context omitted.

Competition keeps service quality high and pricing low - even if you aren't using Grok, the mere existence of Grok keeps pricing for whatever provider you use lower and service faster and more reliable.

That's fair. But my point is from a business POV, why would SpaceX want to invests hundreds of billions of CapEx on a third or fourth frontier model, which cannot compete with chatGPT and claude on the high end, and getting squeezed by open weight models on the low end

[dead]

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#235
post #226
post #31

Earlier quoted context omitted.

Probably can't advance frontier math yet, yeah. But please let us know other places you want to see Grok improve for future models!

Not the model but the app - with Claude I can work on my mac using Cloud environments, leave the office, open Claude on my phone and respond. I can't change model if I have to stop that workflow. Does xAI have plans to do Mac/iOS apps with cloud environments? When can we expect them?

You absolutely can do this; I do it every day in fact. https://apps.apple.com/us/app/cursor/id6767085653

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#236
post #141

Can someone explain to me what's the point of Grok anymore? I don't understand why we need a third or fourth closed frontier model. It is clear that chatGPT has locked down the consumer play, and may be Gemini is there. Claude has enterprise locked up, followed by chatGPT and Gemini. Enterprise switching costs are notoriously high, and even if they switch, they have chatGPT or Gemini to choose from. Beyond that, you…

Its a bet for future world dominance by Elon: millions of robots managed by AI. Base on some observations I think something like that is going in his head.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#237
post #141

Can someone explain to me what's the point of Grok anymore? I don't understand why we need a third or fourth closed frontier model. It is clear that chatGPT has locked down the consumer play, and may be Gemini is there. Claude has enterprise locked up, followed by chatGPT and Gemini. Enterprise switching costs are notoriously high, and even if they switch, they have chatGPT or Gemini to choose from. Beyond that, you…

Why do we need Nissan? Three car companies are plenty.

That’s how physical good work. But software, especially consumer software, works on a winner take all model. That’s why there are very few consumer companies and chatGPT is pretty much the only one after Meta, which was founded in 2004. The reason this happens is because consumer software can scale infinitely as there is zero marginal cost for a new user and there are very high switching costs. A single car company cannot scale to serve every single customer. Because it requires massive CapEx investment. But Google can serve every single search globally because the incremental cost to serve the additional consumer is essentially zero. That’s how these frontier models are eventually going to play out. There will be consolidation and winner take all. It has somewhat happened already with chatGPT taking over consumer and Claude taking over Enterprise. There will probably some long tail open source player, similar to Linux.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#238
post #194

Earlier quoted context omitted.

Are there any projects that track how much usage of each model translates to how much percentage drop in weekly/5hr windows?

Usage? Not exactly. But I tried to make something that can estimate dollars per tokens in actual usage while taking into account multiple factors. https://harness.eveid.com/lazy-harness-cost-simulation

Just curious, was this coded with Claude or Codex? Copy reads very Claude to me but I’m curious if thats an actual pattern or just me

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#240
post #152

Earlier quoted context omitted.

Competition keeps service quality high and pricing low - even if you aren't using Grok, the mere existence of Grok keeps pricing for whatever provider you use lower and service faster and more reliable.

That's fair. But my point is from a business POV, why would SpaceX want to invests hundreds of billions of CapEx on a third or fourth frontier model, which cannot compete with chatGPT and claude on the high end, and getting squeezed by open weight models on the low end

Scamming investors long enough for the founders to cash out.
Post reply on HN