Live data from Hacker News

Viewing profile — stefanbaumann

stefanbaumann

HN member
Joined
Thu, Mar 16, 2023, 5:18 PM UTC
HN karma
143
Public activity
8 items

About stefanbaumann

No profile information was provided.

Recent public activity

  1. comment
    Comment #39115931

    The models presented in the paper are trained on class-conditional ImageNet (where the input is Gaussian noise and one of 1000 classes, e.g., "car") and unconditional FFHQ (where t…

  2. comment
    Comment #39115275

    Not yet, we focused on the architecture for this paper. I totally agree with you though - pixel space is generally less limiting than a latent space for diffusion, so we would expe…

  3. comment
    Comment #39112574

    Both Latent Consistency Models and Adversarial Diffusion Distillation (the method behind SDXL Turbo) are methods that do not depend on any specific properties of the backbone. So, …

  4. comment
    Comment #39112245

    The "input image" is just the noisy sample from the previous timestep, yes. The overall architecture diagram does not explicitly show the conditioning mechanism, which is a small s…

  5. comment
    Comment #39109732

    Thanks a lot! Yeah, the main motivation was trying to find a way to enable transformers to do high-resolution image synthesis: transformers are known to scale well to extreme, mult…

  6. story
  7. comment
    Comment #37475084

    It's already a thing [1]. They also have a project website [2] with some nice videos, although the code hasn't yet been released. [1] https://arxiv.org/abs/2308.09713 [2] https://d…

  8. story