Need for perceptual or adversarial losses in diffusion autoencoders
Determine whether diffusion autoencoders for protein structure reconstruction and generation can obviate the need for perceptual or adversarial losses, and specify the regimes under which such auxiliary losses are necessary or unnecessary to ensure semantic consistency and biophysical plausibility.
References
Part of the original motivation behind diffusion autoencoders was to obviate the need for perceptual and adversarial losses. Whether this is true is still a little unclear; (Sargent et al., 2025) uses perceptual losses during the training process but (Chen et al., 2025) does not.
Training-dynamics analysis further shows that TT-Net's adversarial loss term consistently saturates to a stagnant state across all three noise types, more so than SVD-Net's, while reconstruction quality continues to improve regardless, raising an open question about the adversarial component's contribution that this work identifies but does not resolve.