Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

281–290 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#281
post #136

Earlier quoted context omitted.

True story, for example take Tesla's promise of self driving which was delivered back in 2015.

[flagged]

They defrauded everyone who bought a Tesla on the promise that the Tesla they bought would be fully self-driving sometime in the near future.

Taxis are a different and wholly irrelevant product that doesn’t absolve the fraud that they committed.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#283
post #163
post #102

Earlier quoted context omitted.

If they didn’t constantly reset, they’d be about the same as Anthropic. Right now, I find that Grok offers better value, uses fewer tokens per turn, and makes better code. I haven’t tried Cursor because I don’t want to change editors again, but maybe I should try it…

That is not true; GPT is the most reasoning efficient model family on the market.

The benchmark article we're replying to shows that Grok token usage is at least on par with the latest OpenAI models [1], and significantly cheaper per token:

https://artificialanalysis.ai/models/grok-4-6#token-use

So depending on how you want to define "token efficiency", Grok is either tied with OpenAI, or in the lead.

[1] Though I grant that 4.6 appears to be wordier, on the order of Terra max.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#284
post #226
post #31

Earlier quoted context omitted.

Probably can't advance frontier math yet, yeah. But please let us know other places you want to see Grok improve for future models!

Not the model but the app - with Claude I can work on my mac using Cloud environments, leave the office, open Claude on my phone and respond. I can't change model if I have to stop that workflow. Does xAI have plans to do Mac/iOS apps with cloud environments? When can we expect them?

Options

https://t3.codes/

https://happier.dev/

https://paseo.sh/

All of these also enable multi vendor LLMs.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#285
post #148

Earlier quoted context omitted.

Versus Sam,Dario or the CCP? Im all for running local models but im sor far from being able to pay for a large model hardware setup. My strix halo box is like driving a beaten up vespa when the frontier models are Ferraris.

I probably have the same strix halo box as you. It's slow, although mostly tolerable, but even 128 GB isn't enough to run good models. Where that leaves us is giving money to somebody to get access to frontier models. I don't like any of those guys either, but some appear worse than others. FWIW, I mostly use Anthropic models. And CCP and their distilled models notwithstanding, at least they release open weight model…

Couple of years with some bandwidth improvements and higher RAM capacity might do some wonders

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#286
post #152

Earlier quoted context omitted.

Competition keeps service quality high and pricing low - even if you aren't using Grok, the mere existence of Grok keeps pricing for whatever provider you use lower and service faster and more reliable.

That's fair. But my point is from a business POV, why would SpaceX want to invests hundreds of billions of CapEx on a third or fourth frontier model, which cannot compete with chatGPT and claude on the high end, and getting squeezed by open weight models on the low end

[deleted]

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#288
post #119

Reading the SWE bickering back and fourth in this thread about Claude vs Grok reminds me of IE vs Netscape bickering way back when.

/me eating popcorn from my Lynx term

/me emailing my hot takes and emoticons to newsletters from pine

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#289
post #46

Earlier quoted context omitted.

[flagged]

When you make politics a big part of your identity, you can’t help but to inject it everywhere. It’s glaringly weird and annoying for those of us who are apathetic and just want to talk about tech, but that’s how it has been for a few years.

Get some emotions and care about people

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#290
post #39

Earlier quoted context omitted.

I can respect if you say you hate their guts. Everyone has their worldview. But over moral or ethical stand? You don't have any if you're using Chinese models, or fly Middle East airlines, or countless of other products. Don't delude yourself.

I don’t like what is happening in my country, Elon is a huge part of that and I don’t want to reward it. I’m not a fanatic, but if all other things are equal I can certainly factor social responsibility into the equation. Fuck that rage baiting xenophobic nazi-salute throwing asshole. [edit spelling]

You are confusing cause and effect imho.
Post reply on HN