Qwen3.8 Max now ranked as the best overall model by agentic index
31–40 of 364 posts
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#32I am so excited for Qwen 3.8 27B. It’s a shame how slow prefill (~3-400) is on a strix halo but it’s such a good model for agentic tasks.
How are you running it on a Strix Halo? The weights aren't out yet, are they?
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#33Earlier quoted context omitted.
Indeed, and this Qwen 3.8 max specific page: https://artificialanalysis.ai/models/qwen3-8-max Doesn't have the claim either. Clickbait?
This page has it, scroll to "Intelligence" header (not the highlights one, but second on the page / with black square) and click "Agentic Index"
Even then, this seems a much more marginal win than the headline suggested to me.
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#34Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#35I am so excited for Qwen 3.8 27B. It’s a shame how slow prefill (~3-400) is on a strix halo but it’s such a good model for agentic tasks.
How are you running it on a Strix Halo? The weights aren't out yet, are they?
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#36Strange that the page https://artificialanalysis.ai/agents/coding-agents doesn't even mention "Qwen" once if it's now the "best" according to one of their one index?
But: I've been very impressed by the larger Qwen Models, and a brief try of Kimi also impressed me.
A lingering sense of quality degradation when going deep remains.
But that's not an accusation: they seem to be hitting the compute/quality tradeoff extremely well.
And on-prem capability is simply irreplaceable.
Apart from all the innovations that were driven by the strive for this optimization: quantization, "distilling" (without obvious mad-cows-disease)... I think China was an invaluable player in this progress. Intuitively, I'd even go so far to speculate that LLaMa wouldn't exist without the competition.
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#37Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#38I believe it. It's extremely good at troubleshooting. I gave Qwen and Kimi K3 the same annoying, complicated, intermittent bug to track down. Kimi did a bit better in understanding the existing code, but Qwen built some diagnostic tools and did an excellent statistical analysis on the log data. Qwen got way closer to the truth. I'm very much looking forward to their forthcoming smaller model Qwen 3.8 releases. A vers…
How CLI are you guys using for qwen and kimi?
OpenCode or oh-my-pi might make more sense if you just want a batteries-included agent. You can also make Claude Code work with other models without too much work, but I think that's asking for headaches.
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#39Could someone like Apple be playing the long game - Good Enough(tm) intelligence will eventually fit in our pocket and homes?