Ovi: Twin backbone cross-modal fusion for audio-video generation
1–10 of 122 posts
Re: Ovi: Twin backbone cross-modal fusion for audio-video generation
#2[dead]
Re: Ovi: Twin backbone cross-modal fusion for audio-video generation
#3mindblowing - but still in the uncanny valley. and I guess it's cute that many of the characters live in a world where AI has caused an apocolypse, but is that really the message they want to lead with?
Re: Ovi: Twin backbone cross-modal fusion for audio-video generation
#4Lazyweb: Are these related? If so, how?
Re: Ovi: Twin backbone cross-modal fusion for audio-video generation
#5Lazyweb: Are these related? If so, how? https://news.ycombinator.com/item?id=45603435 https://news.ycombinator.com/item?id=45652726
When a new open weights AI model comes out, opportunists register a domain using its name and start hosting it hoping to make a buck with SEO.
Easier than ever now, as AI-assisted coding tools will build you that generic landing page and basic UI.
Re: Ovi: Twin backbone cross-modal fusion for audio-video generation
#6Seems like the video model is based on Wan2.2.
Lots of activity around Wan lately. It’s nice to see flexible open models make a strong showing against the massively funded closed competitors like OpenAI and Runway.
Re: Ovi: Twin backbone cross-modal fusion for audio-video generation
#7Kinda terrifying. And it can run in 32GB of VRAM? Anyone with a 5090 can start spewing out believable fake videos.
Re: Ovi: Twin backbone cross-modal fusion for audio-video generation
#8How long until we see blockbuster movies produced by a guy in his basement for <$1000?
Re: Ovi: Twin backbone cross-modal fusion for audio-video generation
#9How long until we see blockbuster movies produced by a guy in his basement for <$1000?
I think "blockbuster movie" is a moving target, so it's a bit hard to know
Re: Ovi: Twin backbone cross-modal fusion for audio-video generation
#10How long until we see blockbuster movies produced by a guy in his basement for <$1000?
Soon as the video models can keep characters consistent across scenes. It could take months of prompting to get each scene, but regular movies have long shooting timelines too. If we ever get to instant movie, thought to scene, then movies will die since people will just daydream through the AI.