Understanding the Latent Space of Diffusion Models through the Lens of Riemannian Geometry

07/24/2023
by   Yong-Hyun Park, et al.
0

Despite the success of diffusion models (DMs), we still lack a thorough understanding of their latent space. To understand the latent space 𝐱_t ∈𝒳, we analyze them from a geometrical perspective. Specifically, we utilize the pullback metric to find the local latent basis in 𝒳 and their corresponding local tangent basis in ℋ, the intermediate feature maps of DMs. The discovered latent basis enables unsupervised image editing capability through latent space traversal. We investigate the discovered structure from two perspectives. First, we examine how geometric structure evolves over diffusion timesteps. Through analysis, we show that 1) the model focuses on low-frequency components early in the generative process and attunes to high-frequency details later; 2) At early timesteps, different samples share similar tangent spaces; and 3) The simpler datasets that DMs trained on, the more consistent the tangent space for each timestep. Second, we investigate how the geometric structure changes based on text conditioning in Stable Diffusion. The results show that 1) similar prompts yield comparable tangent spaces; and 2) the model depends less on text conditions in later timesteps. To the best of our knowledge, this paper is the first to present image editing through 𝐱-space traversal and provide thorough analyses of the latent structure of DMs.

READ FULL TEXT

page 6

page 15

page 20

page 21

page 22

page 23

page 24

page 25

research
02/24/2023

Unsupervised Discovery of Semantic Latent Directions in Diffusion Models

Despite the success of diffusion models (DMs), we still lack a thorough ...
research
10/20/2022

Diffusion Models already have a Semantic Latent Space

Diffusion models achieve outstanding generative performance in various d...
research
05/26/2022

Analyzing the Latent Space of GAN through Local Dimension Estimation

The impressive success of style-based GANs (StyleGANs) in high-fidelity ...
research
11/16/2021

Delta-GAN-Encoder: Encoding Semantic Changes for Explicit Image Editing, using Few Synthetic Samples

Understating and controlling generative models' latent space is a comple...
research
06/13/2021

Do Not Escape From the Manifold: Discovering the Local Coordinates on the Latent Space of GANs

In this paper, we propose a method to find local-geometry-aware traversa...
research
06/01/2023

Intriguing Properties of Text-guided Diffusion Models

Text-guided diffusion models (TDMs) are widely applied but can fail unex...
research
05/16/2020

Geodesics in fibered latent spaces: A geometric approach to learning correspondences between conditions

This work introduces a geometric framework and a novel network architect...

Please sign up or login with your details

Forgot password? Click here to reset