Live data from Hacker News

Qwen3.8 Max now ranked as the best overall model by agentic index

artificialanalysis.ai

11–20 of 364 posts

Re: Qwen3.8 Max now ranked as the best overall model by agentic index

#11
post #4

Why does an open weights model cost nearly the same as GPT5.6? $1.14 vs $1.23 on the cost index. Since you can't presumably run this on your own hardware given the model size and hence gain other things like privacy, I don't see any reason to move away from GPT at this rate.

For one thing, providers of open models can't arbitrarily increase their prices without facing competition.

Re: Qwen3.8 Max now ranked as the best overall model by agentic index

#12
post #4

Why does an open weights model cost nearly the same as GPT5.6? $1.14 vs $1.23 on the cost index. Since you can't presumably run this on your own hardware given the model size and hence gain other things like privacy, I don't see any reason to move away from GPT at this rate.

[deleted]

Re: Qwen3.8 Max now ranked as the best overall model by agentic index

#13
post #4

Why does an open weights model cost nearly the same as GPT5.6? $1.14 vs $1.23 on the cost index. Since you can't presumably run this on your own hardware given the model size and hence gain other things like privacy, I don't see any reason to move away from GPT at this rate.

You said it yourself, model size and hardware. Big models cost more (good optimisation reduces things slightly, but they still need the hardware).

Re: Qwen3.8 Max now ranked as the best overall model by agentic index

#14
post #11
post #4

Why does an open weights model cost nearly the same as GPT5.6? $1.14 vs $1.23 on the cost index. Since you can't presumably run this on your own hardware given the model size and hence gain other things like privacy, I don't see any reason to move away from GPT at this rate.

For one thing, providers of open models can't arbitrarily increase their prices without facing competition.

But given the extremely low cost of switching, why wouldn't you use the cheaper one if they're comparable?

Re: Qwen3.8 Max now ranked as the best overall model by agentic index

#15

Strange that the page https://artificialanalysis.ai/agents/coding-agents doesn't even mention "Qwen" once if it's now the "best" according to one of their one index?

Indeed, and this Qwen 3.8 max specific page:

https://artificialanalysis.ai/models/qwen3-8-max

Doesn't have the claim either. Clickbait?

Re: Qwen3.8 Max now ranked as the best overall model by agentic index

#16

Strange that the page https://artificialanalysis.ai/agents/coding-agents doesn't even mention "Qwen" once if it's now the "best" according to one of their one index?

According to those graphs, Grok 4.5 appears to be the most cost-effective model.

Re: Qwen3.8 Max now ranked as the best overall model by agentic index

#17
post #10

Strange that the page https://artificialanalysis.ai/agents/coding-agents doesn't even mention "Qwen" once if it's now the "best" according to one of their one index?

Different benchmarks: > Artificial Analysis Agentic Index: Represents the weighted average of agentic capabilities benchmarks in the Artificial Analysis Intelligence Index (GDPval-AA v2, Tau³-Banking) > Artificial Analysis Coding Agent Index v1.3 incorporates 3 benchmarks: DeepSWE, Terminal-Bench v2, and SWE-Atlas-QnA Qwen3.8 Max is 55.4 on the Agentic Index but hasn't been tested for the Coding Agent Index.

Looks like coding agent is model+harness. There are far fewer models represented on that page. I believe "agentic index" is still the metric to look at for coding performance. I could be wrong about that though.
Post reply on HN