Earlier quoted context omitted.
It’s available in the iOS app (or was for me), both logged in and out.
Is there an iOS app for using Qwen?!
Qwen 3.8
191–200 of 793 posts
Re: Qwen 3.8
#192Earlier quoted context omitted.
I would rather see them releasing 3.7-27B, 3.7-122B or their 3.8 versions. Qwen/QwQ were always about the best available local inference at home.
I know this is a bit cliche but I wonder how much headroom there is in the lower parameter count range. Is there any good reason to believe there is a lot of headroom or there is not? I suppose I'm just wondering if this wave of nearly Fable class models will be runnable on ~$10k worth of hardware at reasonable speeds in the near future.
I would love to see something like a 90B A6B model that is optimized for 128GB machines e.g. strix halo, I haven’t seen anything really targeting the combination of RAM and compute these machines have, but I’m biased because I have one.
Re: Qwen 3.8
#193Earlier quoted context omitted.
Is qwen 3.6 27b the best model you can run locally at the moment? Not that I have the VRAM for it, but just curious.
In my experience, yes. A bit more reliable than gemma for me. I mostly use A3B (35B, mix of experts) though, because it‘s faster, and in the same ballpark intelligence wise as the dense 27B, so it’s the sweetspot for me. I want to try cohere‘s mini code model next, but worried the runtimes aren‘t optimized for that yet.
Re: Qwen 3.8
#194Earlier quoted context omitted.
> How does this explain open weights? They could easily take the same closed route like their American friends Because they are playing the Americans at their own game. What is the first thing an American company would do ? Spread the old American classic FUD ... "you can't used this closed tool because its run by the communists", right ? So you release it as open weights which is a win-win. Global adoption of the mo…
Alibaba isn’t really the Chinese government though, or are you saying Americans will think that ever since Jack Ma was harmonized?
I wasn't saying anything about what Americans would think.
I was saying about what they would inevitably be told by US politicians and by US AI companies.
If you were a sales-rep or marketeer at a US AI company, I bet you would be using the old "evil communists" routine in relation to any closed Chinese model.
I was saying that by releasing as open weights, the company has removed that line of argument.
Clearly I was a bit broad in my use of "the Chinese" when in this case it was, as you say, a Chinese company.
Re: Qwen 3.8
#195So are locally-runnable models frozen at Qwen 3.6 now :/
Is qwen 3.6 27b the best model you can run locally at the moment? Not that I have the VRAM for it, but just curious.
Re: Qwen 3.8
#196Re: Qwen 3.8
#197Earlier quoted context omitted.
My experience was so much different to this, that I have the unfortunate impression that you're shilling. It really was not a capable model, it felt like the old oai models back when we were all excited but couldn't actually trust them even in the littlest ways. What harness were you using, did you do any work to make it better? What was I doing wrong? I just pointed opencode at it, with a pretty simple (large-ish) d…
Have you actually used Opus 4.8 in Claude Code? It takes way too long to do any practical task on higher thinking levels due to over-engineering. And I am not the only one complaining. Lots of people downgrade to Opus 4.6 exactly for this reason. Opus 4.8 training works well for agentic work. Not for code harness. EDIT: ``` stronger on coding and raw capability but can be more argumentative, verbose, and costly. Reli…
Re: Qwen 3.8
#198Earlier quoted context omitted.
Alibaba isn’t really the Chinese government though, or are you saying Americans will think that ever since Jack Ma was harmonized?
> are you saying Americans will think that I wasn't saying anything about what Americans would think. I was saying about what they would inevitably be told by US politicians and by US AI companies. If you were a sales-rep or marketeer at a US AI company, I bet you would be using the old "evil communists" routine in relation to any closed Chinese model. I was saying that by releasing as open weights, the company has r…
Re: Qwen 3.8
#199Earlier quoted context omitted.
> If by Opus you mean Opus 4 and not Opus 4.8, then sure I meant Opus 4.8 which is rather dumb and ineffective in coding harness, especially with higher thinking levels.
Comments like this boggle my mind. The model which everyone else raves about and is wildly successful with legions of programmers virtually demanding access while abandoning ChatGPT and Copilot in droves, is rather dumb? Have you considered that it's more likely that you're doing something wrong?
Re: Qwen 3.8
#200Earlier quoted context omitted.
Is qwen 3.6 27b the best model you can run locally at the moment? Not that I have the VRAM for it, but just curious.
Yes it's between this and Gemma 4 31B which is much slower, but looks like it won't ever get an upgrade. I have to conclude that the MoE variants are unreliable, and MTP sometimes just can't get tricky formatting right.
Not tried it yet but I've seen tests that suggest they've properly fixed the tool calling issues.