Abstract
Recently, diffusion models have made remarkable progress in text-to-image (T2I) generation, synthesizing images with high fidelity and diverse contents. Despite this advancement, latent space smoothness within diffusion models remains largely unexplored. Smooth latent spaces ensure that a perturbation on an input latent corresponds to a steady change in the output image. This property proves beneficial in downstream tasks, including image interpolation, inversion, and editing. In this work, we expose the non-smoothness of diffusion latent spaces by observing noticeable visual fluctuations resulting from minor latent variations. To tackle this issue, we propose Smooth Diffusion, a new category of diffusion models that can be simultaneously high-performing and smooth. Specifically, we introduce Step-wise Variation Regularization to enforce the proportion between the variations of an arbitrary input latent and that of the output image is a constant at any diffusion training step. In addition, we devise an interpolation standard deviation (ISTD) metric to effectively assess the latent space smoothness of a diffusion model. Extensive quantitative and qualitative experiments demonstrate that Smooth Diffusion stands out as a more desirable solution not only in T2I generation but also across various downstream tasks. Smooth Diffusion is implemented as a plug-and-play Smooth-LoRA to work with various community models. Code is available at this https URL.
data:image/s3,"s3://crabby-images/dc661/dc661d3e5ed70abba6258610151de918ef950d08" alt=""
data:image/s3,"s3://crabby-images/d2898/d2898cdb295921588c4392a10600a42b5d23c69d" alt=""
data:image/s3,"s3://crabby-images/d5693/d5693356585f43f7de959c274ad59463c6d3733b" alt=""
data:image/s3,"s3://crabby-images/61e4d/61e4d8fb2a0fa7910f5a4b1ce6de59957b76797d" alt=""
data:image/s3,"s3://crabby-images/07ebe/07ebedaa074e73885f0d6d160f855ff52cb07f84" alt=""
Paper: https://arxiv.org/abs/2312.04410
Code: https://github.com/SHI-Labs/Smooth-Diffusion (coming soon)
Project Page: https://shi-labs.github.io/Smooth-Diffusion/