Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

251–260 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#251

Cursor, since Grok 4.5, has had an incredible deal for frontier level models, their subscription now goes way further than OpenAI or Anthropic. Even on their lower tier plans you can use a lot tokens on their of their first party models (Grok and Composer) and not really run out comparatively. Combine them with an orchestrator and implementor type setup and it goes even further.

>their subscription now goes way further than OpenAI or Anthropic. Until it doesn't... Honestly, this entire OpenAI reset credit fiasco this past week has convinced me to rip off the Codex and Claude Code bandaids and start building my own proper Pi Coding Agent running models that I select and pay for on openrouter. And I am feeling a lot better about it now that I've finally got it working.

> Honestly, this entire OpenAI reset credit fiasco this past week

Huh, what's happened? I'm on the 20x plan and haven't noticed any fiasco, what went down exactly?

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#252

Cursor, since Grok 4.5, has had an incredible deal for frontier level models, their subscription now goes way further than OpenAI or Anthropic. Even on their lower tier plans you can use a lot tokens on their of their first party models (Grok and Composer) and not really run out comparatively. Combine them with an orchestrator and implementor type setup and it goes even further.

I believe they are the only western provider that has Kimi K3 on a subscription plan today as well. I would love to ditch Anthropic and be on Kimi if there were a subsidized plan like that with ZDR

Kimi is expensive . Cursor with subscription is cheaper , grok 4.5 per task paid per tokens ( no subs ) is also cheaper .

If you willing to share to no zdr, meta is waaaaaay cheaper vs Kimi.

With recent offerings from spacex and meta , I hardly imagine why would you pay money to any Chinese vendor it’s not as cheap and it’s not as intelligent neither .

Maybe deepseek is an exception , but it’s only good for narrow use cases that probably goes into modal.com and other gpu + fine tune me easy vendors , not vanilla dumb but cheap model .

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#253
post #136

Earlier quoted context omitted.

True story, for example take Tesla's promise of self driving which was delivered back in 2015.

[flagged]

There’s some number of supposedly unsupervised Teslas actually operating in Austin, but what I’ve seen suggests it’s more like a dozen.

And extreme skepticism is warranted that they’re actually fully unsupervised, given Tesla’s repeated lies about this. They’re likely remotely monitored and operated.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#254
post #238
post #194

Earlier quoted context omitted.

Usage? Not exactly. But I tried to make something that can estimate dollars per tokens in actual usage while taking into account multiple factors. https://harness.eveid.com/lazy-harness-cost-simulation

Just curious, was this coded with Claude or Codex? Copy reads very Claude to me but I’m curious if thats an actual pattern or just me

[deleted]

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#255
post #187

Earlier quoted context omitted.

>> I think they will pull ahead with cheaper tokens similar intelligence they just increased cache read from 0.30 to 0.50 - this has the biggest impact on agentic coding. Elon companies have the most expensive everything: xAI sub: $30 when other starts at $20, pro like sub for $300 where other charge $200. Expensive electric cars, powerwalls, solar roofs when competetive products/better are cheaper.

>Elon companies have the most expensive everything The Model 3 and Model Y became the highest selling EVs of all time because they were the first below $50K to have long-range and be worth buying. Until a few years ago, every other sub-$50K EV absolutely sucked.

highest selling is not same as cheap - apple also have the highest selling iphones among smartphones and they are still pretty much the most expensive (and usually not the best either these days).

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#256
post #49

Earlier quoted context omitted.

once my codex/claude weekly limit was gone, i gave it a try. It was surprisingly good, not dumb in any way, and fast . I now require it as a part of 3-of-3 quorum with any codebase change.

> I now require it as a part of 3-of-3 quorum with any codebase change Say more about this.

macOS Codex app is my main agent (its very polished). When it makes changes or code scans i ask it to run claude -p plus cursor-cli plus grok cli for double checking. This way bias of one model can be overruled by quorum.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#258
post #217
post #14

Seems the cache read pricing almost doubled from $0.30 in Grok 4.5 to $0.50 in Grok 4.6. In my experience in heavy coding sessions most pricing is just cache read and cache write like 80% of my token bill.

didnt the model 3x in size?

no, they said the model is still 1.5T size. Next one 4.7 is supposed to be bigger.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#259

Earlier quoted context omitted.

>their subscription now goes way further than OpenAI or Anthropic. Until it doesn't... Honestly, this entire OpenAI reset credit fiasco this past week has convinced me to rip off the Codex and Claude Code bandaids and start building my own proper Pi Coding Agent running models that I select and pay for on openrouter. And I am feeling a lot better about it now that I've finally got it working.

> Honestly, this entire OpenAI reset credit fiasco this past week Huh, what's happened? I'm on the 20x plan and haven't noticed any fiasco, what went down exactly?

Nothing really, when model usage goes up they do a credit. They did two last week.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#260
post #39
post #18

Earlier quoted context omitted.

I refuse to use that product because of the parent company.

I can respect if you say you hate their guts. Everyone has their worldview. But over moral or ethical stand? You don't have any if you're using Chinese models, or fly Middle East airlines, or countless of other products. Don't delude yourself.

I don’t like what is happening in my country, Elon is a huge part of that and I don’t want to reward it. I’m not a fanatic, but if all other things are equal I can certainly factor social responsibility into the equation. Fuck that rage baiting xenophobic nazi-salute throwing asshole. [edit spelling]
Post reply on HN