Live data from Hacker News

Megaface

exposing.ai

111–114 of 114 posts

Re: Megaface

#111
post #108
post #81

Earlier quoted context omitted.

If there's a specific prompt that makes the model produce such an image (without img2img), why shouldn't it count? The question isn't whether the model produces such things randomly, but rather whether it's capable of producing them in principle, even if it requires a very elaborate prompt.

> If there's a specific prompt that makes the model produce such an image (without img2img), why shouldn't it count? There's a specific prompt which makes a human artist produce such an image without img2img too: "Please draw the Mona Lisa". There - you just did it in your head while reading this comment!

So? Humans aren't copyrightable works. But a bunch of weights that constitutes the model is just data, and that data can very well contain copyrighted works. The question is whether it does in any meaningful sense. And if it can reproduce them verbatim with the right prompt, I don't see how the answer could be "no", anymore so than a password-protected archive of a copyrighted JPEG.

Re: Megaface

#112

Earlier quoted context omitted.

One, the diffusion model's possible output space contains every RGB image ever. But two, it cannot ever possibly contain the original inputs verbatim, because (the size of the model)/(the size of the training set) comes out to be something like 0.2 KB per image. Unless it's an incredible compression algorithm, diffusion necessarily have learned something from the input rather than copy-pasting things, as claimed upth…

> Unless it's an incredible compression algorithm, diffusion necessarily have learned something from the input rather than copy-pasting things, as claimed upthread. Arguably, "learning" and "compression" are the same thing. In this sense, you can view SD as a compression algorithm where the decoder is the model, the compressed file is the prompt + tweakable params, and there aren't any error checks made, so you can f…

Or maybe the compressed file is the model, and prompt + tweakable params is a "path" inside that compressed file?

Re: Megaface

#113

Earlier quoted context omitted.

>Or is it republishing of the original data? If it's publishing _data_ then you're fine under regular copyright as it only protects artistic works and not things like data. You might fall shy of other IP legislation but not copyright. YMMV, this is not legal advice and represents my personal opinion unrelated to my employment.

The "data" here is photographs, which all jurisdictions I'm aware of treat as coprightable.

FWIW in UK it's possible for works to be too generic to attract copyright, or to be slavish reproductions (eg as photos).

ID images have their parameters dictated by technical needs - no smiling, plain background, even lighting, no eyewear, head only, face-on - and so leave no room for artistry.

An ID photo might lack copyright.

I know of no caselaw here (on copyright in ID shots) and am projecting from eg the "red bus" case (Temple Island v New English Tea).

Re: Megaface

#114

Earlier quoted context omitted.

Replace "corporate profit" with "social good", which is what it generally comes from, and then is it still sad? You seem to imply there's something wrong with corporate profit. We as society want and encourage corporate profit because we want the social good that corporations provide and the profit incentivizes them to do it. Profit is a rough measure of how much good they do for people. Profit is like salary for inv…

Corporate profit does not always lead to social good. There are lots of cases, especially when companies scale significantly and over a long enough period of time, where the profit motive leads to a decline in social good which is often seen in the form of negative externalities. > Profit is a rough measure of how much good they do for people. While this is true to a certain degree in many situations it does not capt…

[dead]
Post reply on HN