← Back to context

Comment by efavdb

9 hours ago

For images, latent diffusion only works when using a special decoder that can produce realistic images given relevant samples from the latent space. This is trained like a GAN and doesn’t focus on pixel level error but higher level image features extracted via another network etc. expect that building a decoder like that for their problem may have solved their issues.