Qwen3-VL
qwen.ai
Qwen3-VL
1–10 of 166 posts
Re: Qwen3-VL
#2Re: Qwen3-VL
#3Re: Qwen3-VL
#4Re: Qwen3-VL
#5Re: Qwen3-VL
#6Re: Qwen3-VL
#7https://openrouter.ai/qwen/qwen3-235b-a22b-thinking-2507
Now with this I will use it to identify and caption meal pictures and user pictures for other workflows. Very cool!
Re: Qwen3-VL
#8Re: Qwen3-VL
#9The biggest takeaway is that they claim SOTA for multi-modal stuff even ahead of proprietary models and still released it as open-weights. My first tests suggest this might actually be true, will continue testing. Wow
Doesn't seem to be far ahead of existing proprietary implementations. But it's still good that someone's willing to push that far and release the results. Getting multimodal input to work even this well is not at all easy.
Re: Qwen3-VL
#10Cool! Pity they are not releasing a smaller A3B MoE model
Relevant comparison is on page 15: https://arxiv.org/abs/2509.17765