Edit: The larger models have 128k context length. 32k thinking comes from the chart which looks like it's for the 235B, so not full length.
Qwen3: Think deeper, act faster
11–20 of 412 posts
Re: Qwen3: Think deeper, act faster
#12They have got pretty good documentation too[1]. And Looks like we have day 1 support for all major inference stacks, plus so many size choices. Quants are also up because they have already worked with many community quant makers. Not even going into performance, need to test first. But what a stellar release just for attention to all these peripheral details alone. This should be the standard for major release, inste…
they have already worked with many community quant makers I’m curious, who are the community quant makers?
Re: Qwen3: Think deeper, act faster
#13Out of all the Qwen3 models on Hugging Face, it's the most downloaded/hearted. https://huggingface.co/collections/Qwen/qwen3-67dd247413f0e2...
Re: Qwen3: Think deeper, act faster
#14Re: Qwen3: Think deeper, act faster
#15They have got pretty good documentation too[1]. And Looks like we have day 1 support for all major inference stacks, plus so many size choices. Quants are also up because they have already worked with many community quant makers. Not even going into performance, need to test first. But what a stellar release just for attention to all these peripheral details alone. This should be the standard for major release, inste…
they have already worked with many community quant makers I’m curious, who are the community quant makers?
Re: Qwen3: Think deeper, act faster
#16They have got pretty good documentation too[1]. And Looks like we have day 1 support for all major inference stacks, plus so many size choices. Quants are also up because they have already worked with many community quant makers. Not even going into performance, need to test first. But what a stellar release just for attention to all these peripheral details alone. This should be the standard for major release, inste…
Well, the link to huggingface is broken at the moment.
The space loads eventually as well; might just be that HF is under a lot of load.
Re: Qwen3: Think deeper, act faster
#17We're really getting close to the point where local models are good enough to handle practically every task that most people need to get done.
Re: Qwen3: Think deeper, act faster
#18A 0.6B LLM with a 32k context window is interesting, even if it was trained using only distillation (which is not ideal as it misses nuance). That would be a fun base model for fine-tuning. Out of all the Qwen3 models on Hugging Face, it's the most downloaded/hearted. https://huggingface.co/collections/Qwen/qwen3-67dd247413f0e2...
my concern on these models though unfortunately is it seems like architectures very a bit so idk how it'll work