I got access to the preview, here's what it gave me for "A pelican riding a bicycle along a coastal path overlooking a harbor" - this video has all four versions shown: https://static.simonwillison.net/static/2024/pelicans-on-bic... Of the four two were a pelican riding a bicycle. One was a pelican just running along the road, one was a pelican perched on a stationary bicycle, and one had the pelican wearing a weird…
There's another important contender in the space: Hunyuan model from Tencent My company (Nim) is hosting Hunyuan model, so here's a quick test (first attempt) at "pelican riding a bycicle" via Hunyuan on Nim: https://nim.video/explore/OGs4EM3MIpW8 I think it's as good, if not better than Sora / Veo
Veo 2: Our video generation model
191–200 of 342 posts
Re: Veo 2: Our video generation model
#192Earlier quoted context omitted.
Everyone has access to YouTube. It’s safe to assume that Sora was trained on it as well.
All you can eat? Surely they charge a lot for that, at least. And how would you even find all the videos?
Re: Veo 2: Our video generation model
#193This might be a dumb question to ask, but what exactly is this useful for? B-Roll for YouTube videos? I'm not sure why so much effort is being put into something like this when the applications are so limited.
Re: Veo 2: Our video generation model
#194It's interesting they host these videos on YouTube, cause it signals they're fine with AI generated content. I wonder if Google forgets that the creators themselves are what makes YouTube interesting for viewers.
Re: Veo 2: Our video generation model
#195It’s telling that safety and responsibility gets so much fluff words, technical details are fairly extensive, but no mention of the training data? It’s clearly relevant for both performance and ethical discussions. Maybe it’s just me who couldn’t find it, (the website barely works at all on FF iOS)..
Re: Veo 2: Our video generation model
#196My theory as to why all the bigtech companies are investing so much money in video generation models is simple: they are trying to eliminate the threat of influencers/content creators to their ad revenue. Think about it, almost everyone I know rarely clicks on ads or buys from ads anymore. On the other hand, a lot of people including myself look into buying something advertised implicitly or explicitly by content cre…
Re: Veo 2: Our video generation model
#197Earlier quoted context omitted.
There's another important contender in the space: Hunyuan model from Tencent My company (Nim) is hosting Hunyuan model, so here's a quick test (first attempt) at "pelican riding a bycicle" via Hunyuan on Nim: https://nim.video/explore/OGs4EM3MIpW8 I think it's as good, if not better than Sora / Veo
Reddit says it is much better than Sora. Are you hosting the full version of Nunyuan? (Your video looks great.)
Re: Veo 2: Our video generation model
#198Earlier quoted context omitted.
I mean, I have trained myself on Youtube. Why can't a silicon being train itself on Youtube as well?
When a company trains an AI model on something, and then that company sells access to the ai model, the company, not the ai model, is the being violating copyright. If Jimmy makes an android in his garage and gives it free will, then it trains itself on youtube, i doubt anyone would have an issue.
Re: Veo 2: Our video generation model
#199Superficially impressive but what is the actual use case of the present state of the art? It makes 10-second demos, fine. But can a producer get a second shot of the same scene and the same characters, with visual continuity? Or a third, etc? In other words, can it be used to create a coherent movie --even a 60-second commercial -- with multiple shots having continuity of faces, backgrounds, and lighting? This quote…
Re: Veo 2: Our video generation model
#200Superficially impressive but what is the actual use case of the present state of the art? It makes 10-second demos, fine. But can a producer get a second shot of the same scene and the same characters, with visual continuity? Or a third, etc? In other words, can it be used to create a coherent movie --even a 60-second commercial -- with multiple shots having continuity of faces, backgrounds, and lighting? This quote…