Live data from Hacker News

Stable Video Diffusion

stability.ai

251–260 of 316 posts

Re: Stable Video Diffusion

#251

It's funny that still don't really have video wallpapers on most devices (I'm only aware of Wallpaper Engine on Windows)

I had a video wallpaper on my Motorola Droid back in 2010.

and a battery life of...?

I do wonder if there have been any codec studies that measure power usage with respect to RAM

Re: Stable Video Diffusion

#252
How much longer will it be until we can play "video games" which consist of user-input streamed to an AI that generates video output and streams it to the player's screen?

Re: Stable Video Diffusion

#253

Earlier quoted context omitted.

But surely you wouldn't try to emit that format directly, but rather some higher level scene description? Or even just a set of instructions for how to manipulate the UI to create the imagined scene?

It sure feels weird to me as well, that GenAI is always supposed to be end-to-end with everything done inside NN blackbox. No one seems to be doing image output as SVG or .ai.

There is a fundamental disconnect between industry and academia here.

Re: Stable Video Diffusion

#254

Earlier quoted context omitted.

Maybe something like OpenSCAD is a good middle ground. Procedural code-like format for specifying 3D objects that can then be converted and imported in Blender.

I tried all the AI stuff that I could on OpenSCAD. While it generates a lot of code that initially makes sense, when you use the code, you get a jumbled block.

This. I think problem is that the LLMs really struggle with 3d scene understanding, so what you would need to do is generate code that generates code.

But also I suspect there just isn't that much openscad code in the training data, and the semantics are different enough to python or any of the other languages that are well-represented that it struggles.

Re: Stable Video Diffusion

#255

It's funny that still don't really have video wallpapers on most devices (I'm only aware of Wallpaper Engine on Windows)

Mplayer/MPV used to be able to play videos in the X root window like a wallpaper. No idea if it still works nowadays.

Re: Stable Video Diffusion

#256

Can this be used for porn?

The answer to that question is always "yes", regardless what "this" is.

Diffusion models for moving images are already used to a limited extent for this. And I'm sure it will be the use case, not just an edge case.

Re: Stable Video Diffusion

#257

Earlier quoted context omitted.

But surely you wouldn't try to emit that format directly, but rather some higher level scene description? Or even just a set of instructions for how to manipulate the UI to create the imagined scene?

It sure feels weird to me as well, that GenAI is always supposed to be end-to-end with everything done inside NN blackbox. No one seems to be doing image output as SVG or .ai.

Imo the thinking is that whenever humans have tried to pre-process or feature-engineer a solution or tried to find clever priors in the past, massive self-supervised-learning enabled, coarsely architected, data-crunching NNs got better results in the end. So, many researchers / industry data scientists may just be disinclined to put effort into something that is doomed to be irrelevant in a few years. (And, of course, with every abstraction you will lose some information that may bear more importance than initially thought)

Re: Stable Video Diffusion

#258

In the video towards the bottom of the page, there are two birds (blue jays), but in the background there are two identical buildings (which look a lot like the CN Tower). CN Tower is the main landmark of Toronto, whose baseball team happens to be the Blue Jays. It's located near the main sportsball stadium downtown. I vaguely understand how text-to-image works, and so it makes sense that the vector space for "blue j…

> Has anyone come across a solution where model can iterate (eg, with prompts like "move the bicycle to the left side of the photo")? It feels like we're close. I feel like we're close too, but for another reason. For although I love SD and these video examples are great... It's a flawed method: they never get lighting correctly and there are many incoherent things just about everywhere. Any 3D artist or photographer…

Not that I’m against the described 3d way, but personally I don’t care about light and shadows until it’s so bad that I do. This obsession with realism is irrational in video games. In real life people don’t understand why light works like this or like that. We just accept it. And if you ask someone to paint how it should work, the result is rarely physical but acceptable. It literally doesn’t matter until it’s very bad.

Re: Stable Video Diffusion

#259
post #184

Earlier quoted context omitted.

Why not? Two willing parties can agree to bind themselves to all kinds of obligations in a contract as long as they're not explicitly illegal. Copyleft is an example of someone successfully inventing a copyright-like right by bootstrapping off existing copyright with a specially engineered contract.

There are a few problems: 1) You and I invent our own private "copyright" for data (which is not copyrightable) 2) Everything is fine until my wife walks up to my computer and makes a copy of the data. She's not bound by our private "copyright." She doesn't even know it exists, and shares the data with her bestie. And... our private pseudo-copyright is dead. Also: Licenses are not the same as contracts. There are tim…

> my wife walks up to my computer and makes a copy of the data

As you agreed to in our contract, you now need to compensate me for the damage caused by your failure to prevent unauthorized third-party access. Of course you're free to attempt to recover the sum you have to pay me from your wife.

> The output of a program is rarely copyrightable by the author (as opposed to the user).

The author of the program can make it a condition of letting the user use the program that the user has to assign all copyright to the author of the program, kind of like "By uploading any User Content you hereby grant and will grant Y Combinator and its affiliated companies a nonexclusive, worldwide, royalty free, fully paid up, transferable, sublicensable, perpetual, irrevocable license to copy, display, upload, perform, distribute, store, modify and otherwise use your User Content for any Y Combinator-related purpose in any form, medium or technology now known or later developed." https://www.ycombinator.com/legal/

Re: Stable Video Diffusion

#260

Earlier quoted context omitted.

> For although I love SD and these video examples are great... It's a flawed method: they never get lighting correctly and there are many incoherent things just about everywhere. Any 3D artist or photographer can immediately spot that. The question is whether the 99% of the audience would even care...

Of course they would. The internet spent a solid month laughing at the Sonic the Hedgehog movie because Sonic had weird-looking teeth.

No they laughed at it because it looked awful in every single way
Post reply on HN