I'm still puzzled as to how these "non-commercial" model licenses are supposed to be enforceable. Software licenses govern the redistribution of the software , not products produced with it. An image isn't GPL'd because it was produced with GIMP.
Nobody claimed otherwise?
Stable Video Diffusion
31–40 of 316 posts
Re: Stable Video Diffusion
#32Looks like I'm still good for my bet with some friends that before 2028 a team of 5-10 people will create a blockbuster style movie that today costs 100+ million USD on a shoestring budget and we won't be able to tell.
Re: Stable Video Diffusion
#33In the video towards the bottom of the page, there are two birds (blue jays), but in the background there are two identical buildings (which look a lot like the CN Tower). CN Tower is the main landmark of Toronto, whose baseball team happens to be the Blue Jays. It's located near the main sportsball stadium downtown. I vaguely understand how text-to-image works, and so it makes sense that the vector space for "blue j…
Re: Stable Video Diffusion
#34It's crazy to see this level of progress in just a bit over half a year.
[1]: https://epiccoleman.com/posts/2023-03-05-deforum-stable-diff...
Re: Stable Video Diffusion
#35I admit I'm ignorant about these model's inner workings, but I don't understand why text is the chosen input format for these models. It was the same for image generation, where one needed to produce text prompts to create the image, and stuff like img2img and Controlnet that allowed things like controlling poses and inpainting, or having multiple prompts with masks controlling which part of the image is influenced b…
Re: Stable Video Diffusion
#36I'm still puzzled as to how these "non-commercial" model licenses are supposed to be enforceable. Software licenses govern the redistribution of the software , not products produced with it. An image isn't GPL'd because it was produced with GIMP.
Re: Stable Video Diffusion
#37The rate of progress in ML this past year has been breath taking. I can’t wait to see what people do with this once controlnet is properly adapted to video. Generating videos from scratch is cool, but the real utility of this will be the temporal consistency. Getting stable video out of stable diffusion typically involves lots of manual post processing to remove flicker.
Re: Stable Video Diffusion
#38Can this be used for porn?
Re: Stable Video Diffusion
#39Can't wait for these things to not suck
Re: Stable Video Diffusion
#40A seemingly off topic question, but with enough compute and optimization, could you eventually simulate “reality”? Like, at this point, what are the technical counters to the assertion that our world is a simulation?
Let's steel-man — you mean 3D VR. Let's stipulate there's a headset today that renders 3D visually indistinguishable from reality. We're still short the other 4 senses
Much like faith, there's always a way to sort of escape the traps here and say "can you PROVE this is base reality"
The general technical argument against "brain in a vat being stimulated" would be the computation expense of doing such, but you can also write that off with the equivalent of foveated rendering but for all senses / entities