Live data from Hacker News

AI real-time human full-body photo generator

generated.photos

121–130 of 421 posts

Re: AI real-time human full-body photo generator

#121
post #62

Oh yeah, totally ready for prime time, hyper realistic, SFW filter works great, not at all hallucinations /massive_sarcasm Actually NSFW, not safe for sanity. That's...not how body parts work: https://generated.photos/human-generator/64e644f39c8c0400108... Prompt was "young woman with tattoos in miniskirt" really nothing crazy there. But perhaps the latent space with that particular pose is particularly raunchy.

Looks like something from a hellraiser movie

[deleted]

Re: AI real-time human full-body photo generator

#122
NSFW? Also.. WTF with their detection algorithm, this is easily abusable. This was the first image I was prompted with https://generated.photos/human-generator/64e65d5a8448b8000b5... I have not changed any of the parameters, they were automatically generated on /new

Re: AI real-time human full-body photo generator

#123
post #115

Earlier quoted context omitted.

> The copyright office has stated they wont grant monopoly privilege for ai generated art. No, they haven't. They've said that if the only human input is a text prompt, then it lacks the required human creativity to be eligible for copyright protection.

Not trying to be combative, but I don't see the difference?

Real AI imagegen workflows very often have more input from the human creating the image than a text prompt.

Re: AI real-time human full-body photo generator

#124
post #24

Why is it impossible to generate a male model wearing anything other than rolled denim jean shorts? I've tried things like "long pants" or "ankle-length pants," but I cannot get it to stop putting them all in denim shorts!

> Why is it impossible to generate a male model wearing anything other than rolled denim jean shorts?

Its not, I got rolled-but-long white denim pants, white shirt, white tie, and white jacket, with white deck-ish shoes but selecting “Formal” on the clothing tab and entering “Clothing: white-tie formalwear”.

But, yeah, there is a definite denim bias.

Re: AI real-time human full-body photo generator

#125
post #62

Oh yeah, totally ready for prime time, hyper realistic, SFW filter works great, not at all hallucinations /massive_sarcasm Actually NSFW, not safe for sanity. That's...not how body parts work: https://generated.photos/human-generator/64e644f39c8c0400108... Prompt was "young woman with tattoos in miniskirt" really nothing crazy there. But perhaps the latent space with that particular pose is particularly raunchy.

There is a SFW filter?

I just let it generate a random woman with no prompt, and it gave me a pretty good result, except there is a mask on the face and literally bloody nude boobs; https://generated.photos/human-generator/64d67874568faa0007a...

edit; I just realized it put in a default prompt

Re: AI real-time human full-body photo generator

#126
post #86
post #9

If you're wondering how it's so fast and cheap and they can generate variants so easily, it's because they're using GANs (see the footer). GANs are way faster than diffusion models because they generate the image in a single forward pass and their true latent space encoding makes editing a breeze. (And if you're wondering how it can look so good when 'everyone knows GANs don't work because they're too unstable', a wi…

>they're too unstable', a widespread myth >See for example BigGAN I remember when you try training a BigGAN model on anime images, the quality was bad. Now look at this example, one single GPU, 1.5M images with a diffusion model: https://medium.com/@enryu9000/anifusion-diffusion-models-for... ,the difference in quality is absurd, you can say this or that is not true but the quality speak for itself, obtaining good qu…

> I remember when you try training a BigGAN model on anime images, the quality was bad

Because there was a bug in the code, in a part unrelated to the GAN itself.

> the difference in quality is absurd

Yes, it does help to train on anime with code that isn't buggy. (BTW, Skylion was getting good results with GANs on anime similarly restricted to centered figures like those samples, he just refuses to ever publish anything.)

Re: AI real-time human full-body photo generator

#127
post #9

If you're wondering how it's so fast and cheap and they can generate variants so easily, it's because they're using GANs (see the footer). GANs are way faster than diffusion models because they generate the image in a single forward pass and their true latent space encoding makes editing a breeze. (And if you're wondering how it can look so good when 'everyone knows GANs don't work because they're too unstable', a wi…

> everyone knows GANs don't work because they're too unstable Is that a wide spread myth? I thought it’s widely accepted that GAN is really good generating these artificial pictures (it’s what started DeepFake after all) when you know your model’s “button”. Similar to how this uses GAN since they have a model “boundary condition”. While humans are diverse, we have a set of repeatable features (two legs, two arms, etc…

It is very widespread. You will see people in this very thread dismissing GANs as fundamentally failed, and hotly objecting to any kind of parity, even if they have to fall back to 'well ok GANs do scale, but they're more complicated'. I also have some representative quotes in my linked draft essay from various papers & DL Twitter discussions. (Another way to put it would be: when was the last time you saw someone besides me asserting that GANs can scale to high-quality general images and are not dramatically inferior to diffusion? I rest my case.)

Re: AI real-time human full-body photo generator

#128

If you refuse their tracking and marketing cookies it redirects you to google.com. Classy.

I'm surprised browsers don't offer something like Docker so that each site is isolated to its own virtual environment.

private/incognito window?

Re: AI real-time human full-body photo generator

#129
post #68
post #9

If you're wondering how it's so fast and cheap and they can generate variants so easily, it's because they're using GANs (see the footer). GANs are way faster than diffusion models because they generate the image in a single forward pass and their true latent space encoding makes editing a breeze. (And if you're wondering how it can look so good when 'everyone knows GANs don't work because they're too unstable', a wi…

You’re overstating the simplicity of a scaling a GAN well. GigaGAN is the best quality out of those and requires 7 loss functions and is incredibly complicated. Sure GANs can scale, but Diffusion models are drastically easier to scale.

No, I'm not. BigGAN did fine on scaling up to JFT-300M with basically no changes beyond model size and a simple architecture. This is also what we were observing, even with a buggy BigGAN implementation. GigaGAN is the best quality, but that's mostly because it's also the biggest; as Table 1 shows most of the gains come from various kinds of additional scaling. (And this is moving the goalposts from the usual assertion that "GANs can't scale" to "they're harder to scale"; note the self-fulfilling nature of such assertions. Considering how there is next to no GAN scaling research, these results are remarkable and show how much low-hanging fruit there is.)

Diffusion models are only 'drastically easier to scale' because researchers have spent the past 3 years researching pretty much nothing but diffusion models, discovering all sorts of subtle issues with them and how to make them scale, which is why it took them so long to become SOTA, and why massive architectural sweeps like https://arxiv.org/abs/2206.00364#nvidia were necessary to discover what makes them 'easier to scale'. If this level of brute force and moon math is 'easy', lord save us from any architecture which is 'hard'!

Re: AI real-time human full-body photo generator

#130

If you refuse their tracking and marketing cookies it redirects you to google.com. Classy.

That violates EU law and you can absolutely get a fine for this behaviour. As a digital service offerer you can ask the user for permission to track non-essential information about the user, but your service should work the same, without regard for if that user says yes or no.

If this service is hell bent on raping your privacy, they will have to limit their offerings to mostly those living in dictatorships and immature democracies.

Post reply on HN