Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
61–70 of 86 posts
Re: Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
#62Hope this on day will be used for auto-tagging all video assets with time codes. The dream of being able to search for running horse and find a clip containing a running horse at 4m42s in one of thousands of clips.
you can do that with Morphik already :) We use an embedding model that processes videos and allows you to perform RAG on them.
Re: Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
#63Earlier quoted context omitted.
No mention of palantir?
Palantir's just the new guy on the block: https://en.wikipedia.org/wiki/Sentient_(intelligence_analysi...
Re: Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
#64Hope this on day will be used for auto-tagging all video assets with time codes. The dream of being able to search for running horse and find a clip containing a running horse at 4m42s in one of thousands of clips.
you can do that with Morphik already :) We use an embedding model that processes videos and allows you to perform RAG on them.
Re: Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
#65Re: Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
#66Re: Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
#67For anyone using Qwen3-VL: where are you running it? I had tons of reliability problems with Qwen3-VL inference providers on OpenRouter — based on uptime graphs I wasn’t alone. But when it worked, Qwen3-VL was pack-leading good at AI Vision stuff.
Re: Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
#68I think Gemini analyzes the transcription.
Can I do the same for free with Qwen3?
Re: Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
#69Hope this on day will be used for auto-tagging all video assets with time codes. The dream of being able to search for running horse and find a clip containing a running horse at 4m42s in one of thousands of clips.
Disclaimer: co-founder
Re: Qwen3-VL can scan two-hour videos and pinpoint nearly every detail
#70The github spells it out much better: https://github.com/QwenLM/Qwen3-VL?tab=readme-ov-file#cookbo...