Live data from Hacker News

Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

artificialanalysis.ai

461–470 of 472 posts

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#461

Earlier quoted context omitted.

> With this comes the consideration about the tradeoffs of "produce mediocre but large body of works" vs "produce high quality but small body of works". Not really. The AI will output the same sort of code at a certain skill level regardless of speed, it's not a human so the above is a false dichotomy. Also, it's sort of strange that you're saying it's inventing work out of nothing, you've never heard of a backlog? I…

> The AI will output the same sort of code at a certain skill level regardless of speed, it's not a human so the above is a false dichotomy. It'll output the same code given the same prompts yes, but you don't just accept whatever it puts out, it requires iterations before it's actually ready to be committed as none of the agents write perfect code on their first try. So, it's not a "false dichotomy", I'm just lookin…

Ah, I see the confusion now. You are assuming one looks at the code at all and adjusts it to fit whatever style guide is needed. The parent you were initially talking to is talking about vibe coding, where whatever the agent spits out is accepted on the first try as long as it works for a given feature. Therefore the more usage one has the more features one can build.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#462
post #428

Earlier quoted context omitted.

If you check https://opencode.ai/go and fold out the what models do you have question you can see a clear list of models included.

i just checked yet again, and i can't see how it satisfies my ask: which models to i have to turn on Chinese providers for or not?

You don’t have to turn anything on to use the providers in that list. They work out of the box.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#463

Earlier quoted context omitted.

What has the elon-verse been spewing out that has you in such an altered state?

What has reddit ben spewing out that has you in such an altered state, divorced from reality?

What is your fixation and preoccupation with that site? Also, you sound unwell...

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#464
post #462

Earlier quoted context omitted.

i just checked yet again, and i can't see how it satisfies my ask: which models to i have to turn on Chinese providers for or not?

You don’t have to turn anything on to use the providers in that list. They work out of the box.

I'm pretty sick of having to address incorrect points masquerading as being well-informed, helpful, or knowing. Please: stop it! Stop it at once. You are wrong. Stop talking like you know anything, because you have only mislead and provided incomplete or outright bad information. Stop hallucinating into this thread! It's rude and bad.

On the OpenCode Go workspace there is a huge box:

  Providers
  Control which providers are used for routing.

  Enable models hosted in China 
If you turn this off, DeepSeek Flash & Pro both stop working:

  Error: Provider request failed with HTTP 403: The latest version of this model is only available hosted in China and requires explicit opt in: https://opencode.ai/workspace/wrk_msh_is_a_liar_stop_bullshitting_please_be_real_example_url/go
Kimi K3, MiMo V2.5 Pro, MiniMax M3, Qwen3.8 Max all work (remarkably!) without the box ticked! That was a strong surprise for me.

My request stands: I would like to have up-front information on what models only have Chinese providers. This information is not available except by buying a plan and experimenting, currently. And I had to try each one to find out, which I wish could be avoided. I want to end this thread where I started it (now that I have done the work to cut through the din and noise and misinformation), with my original request: please OpenCode Go make the information about which providers have Chinese-only/non-China hosting readily available on your website.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#465
post #19

I have never met a single human being who uses Grok for coding

I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business. Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elo…

This a a very mature, stable perspective of the world we live in. Everyone wont agree with me all the time, I don't have to either. And thats ok.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#466
post #331

Earlier quoted context omitted.

>Elon companies have the most expensive everything The Model 3 and Model Y became the highest selling EVs of all time because they were the first below $50K to have long-range and be worth buying. Until a few years ago, every other sub-$50K EV absolutely sucked.

> The Model 3 and Model Y became the highest selling EVs of all time because they were the first below $50K to have long-range and be worth buying. And then Musk totally abandoned Tesla's original brilliant game plan of using the luxury models to find actual low-cost models (which $50K is not), and completely ceded the future EV market to Chinese companies that understand how to make a better car than Tesla for less…

>to find actual low-cost models (which $50K is not)

Model 3 starts at $37K, and Model Y at $40K.

They are #1 and #2 best-selling EVs of all time! So clearly they are affordable.

>BYD sells more cars than Tesla.

Sure, but only by 20%, and only because it also sells low-cost EVs that have less range than anything Tesla sells.

If you exclude BYD EVs with lower range than any Tesla, Tesla outsells BYD by 50%.

Even if you don't, second place (which Tesla still holds) is really not "completely ceded"!

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#467

Earlier quoted context omitted.

Who said it's easy? xAI staff are putting in 80+ hour weeks and building datacenters faster than anyone.

their entire founding team left recently for one. they are the weakest in mission and dont particularly pay much either

If their entire founding team left and now Grok has caught up to the frontier, that says more about the founding team than xAI.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#468

I've been using grok 4.5 with grok build soon after it came out and dropped claude. primarily for personal code. It communicates better. While that might not sound like a big deal it is. It doesn't give me a wall of text, tells me what I need to know and I'll make the actual decisions. It is very quick as well which means the sessions are far more interactive, I'll be steering it more. I sometimes cross check with co…

Out of interest, do Musk's politics impact your decision on whether or not to use Grok? I'd be interested to know where folks lie on the (Agree / Disagree) and (Use / Don't use) axes.

[dead]

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#469
post #462

Earlier quoted context omitted.

You don’t have to turn anything on to use the providers in that list. They work out of the box.

I'm pretty sick of having to address incorrect points masquerading as being well-informed, helpful, or knowing. Please: stop it! Stop it at once. You are wrong. Stop talking like you know anything, because you have only mislead and provided incomplete or outright bad information. Stop hallucinating into this thread! It's rude and bad. On the OpenCode Go workspace there is a huge box: Providers Control which providers…

My answer is still correct. You asked if you had to turn on Chinese models to use this. And no you don’t, the models are available by default. If they work when you manually turn off Chinese models is a different question.

Re: Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index

#470
post #106

Earlier quoted context omitted.

But can you trust any number or metric coming out of SpaceX given everything? Also you mean their own compute like the illegal data centre turbines ? https://www.theguardian.com/technology/2026/jan/15/elon-musk...

Now do Anthropic's copyright fines.

Crazy that posters here don't know the difference between a "fine," a "judgment," and a voluntary settlement. Anthropic didn't even train on the pirated libraries, they said. Still, they made copies. Probably the biggest legal blunder I've seen given the size of the settlement. Alas, it's not a fine.
Post reply on HN