Earlier quoted context omitted.
Which discord if its open to the public? I was on one woth kath in 2021 and loved her insights, would love to again
Same; a good ML focused discord would be great. Training ViTs all day is lonely work. I'm mostly locked into skimming the "Research" channels of image generation discords. LAION used to be decent with a good amount of interesting discussion, but it seems to have devolved into toxicity in the last year.
Direct pixel-space megapixel image generation with diffusion models
41–50 of 50 posts
Re: Direct pixel-space megapixel image generation with diffusion models
#42I'm one of the authors; happy to answer questions. this arch is of course nice for high-resolution synthesis, but there's some other cool stuff worth mentioning.. activations are small! so you can enjoy bigger batch sizes. this is due to the 4x patching we do on the ingress to the model, and the effectiveness of neighbourhood attention in joining patches at the seams. the model's inductive biases are pretty different…
Re: Direct pixel-space megapixel image generation with diffusion models
#43I'm one of the authors; happy to answer questions. this arch is of course nice for high-resolution synthesis, but there's some other cool stuff worth mentioning.. activations are small! so you can enjoy bigger batch sizes. this is due to the 4x patching we do on the ingress to the model, and the effectiveness of neighbourhood attention in joining patches at the seams. the model's inductive biases are pretty different…
Re: Direct pixel-space megapixel image generation with diffusion models
#44I'm one of the authors; happy to answer questions. this arch is of course nice for high-resolution synthesis, but there's some other cool stuff worth mentioning.. activations are small! so you can enjoy bigger batch sizes. this is due to the 4x patching we do on the ingress to the model, and the effectiveness of neighbourhood attention in joining patches at the seams. the model's inductive biases are pretty different…
Did you do any inpainting experiments? I can imagine a pixel-space diffusion model to be better at it than one with a latent auto-encoder.
Re: Direct pixel-space megapixel image generation with diffusion models
#45I'm one of the authors; happy to answer questions. this arch is of course nice for high-resolution synthesis, but there's some other cool stuff worth mentioning.. activations are small! so you can enjoy bigger batch sizes. this is due to the 4x patching we do on the ingress to the model, and the effectiveness of neighbourhood attention in joining patches at the seams. the model's inductive biases are pretty different…
I see your headline speed comparison is to "Pixel-space DiT-B/4" - but how does your model compare to the likes of SDXL? I gather they spent $$$$$$ on training etc, so I'd understand if direct comparisons don't make sense.
And do you have any results on things that are traditionally challenging for generative AI, like clocks and mirrors?
Re: Direct pixel-space megapixel image generation with diffusion models
#46Re: Direct pixel-space megapixel image generation with diffusion models
#47Re: Direct pixel-space megapixel image generation with diffusion models
#48Can I ask one basic thing --> From what are the images generated?
Re: Direct pixel-space megapixel image generation with diffusion models
#49Can I ask one basic thing --> From what are the images generated?
Re: Direct pixel-space megapixel image generation with diffusion models
#50Earlier quoted context omitted.
Making a bank account work for you is a hard discipline and requires budgeting and the like. True, not all of us can "hack" it, but that doesn't mean that with some community classes and help you'll be able to use your bank account well! Why do I feel like a chatbot with this message.
You're actually talking to a bot, in this particular case. 12 minutes old with -2 karma. :berk: