Diversifying Semantic Image Synthesis and Editing via Class- and Layer-wise VAEs

06/25/2021
by   Yuki Endo, et al.
18

Semantic image synthesis is a process for generating photorealistic images from a single semantic mask. To enrich the diversity of multimodal image synthesis, previous methods have controlled the global appearance of an output image by learning a single latent space. However, a single latent code is often insufficient for capturing various object styles because object appearance depends on multiple factors. To handle individual factors that determine object styles, we propose a class- and layer-wise extension to the variational autoencoder (VAE) framework that allows flexible control over each object class at the local to global levels by learning multiple latent spaces. Furthermore, we demonstrate that our method generates images that are both plausible and more diverse compared to state-of-the-art methods via extensive experiments with real and synthetic datasets inthree different domains. We also show that our method enables a wide range of applications in image synthesis and editing tasks.

READ FULL TEXT

page 1

page 2

page 3

page 4

page 5

page 9

page 10

page 12

research
11/04/2021

EditGAN: High-Precision Semantic Image Editing

Generative adversarial networks (GANs) have recently found applications ...
research
12/04/2021

SemanticStyleGAN: Learning Compositional Generative Priors for Controllable Image Synthesis and Editing

Recent studies have shown that StyleGANs provide promising prior models ...
research
06/03/2020

Nested Scale Editing for Conditional Image Synthesis

We propose an image synthesis approach that provides stratified navigati...
research
11/29/2018

Diverse Image Synthesis from Semantic Layouts via Conditional IMLE

Most existing methods for conditional image synthesis are only able to g...
research
03/11/2021

Diverse Semantic Image Synthesis via Probability Distribution Modeling

Semantic image synthesis, translating semantic layouts to photo-realisti...
research
07/11/2023

Automatic Generation of Semantic Parts for Face Image Synthesis

Semantic image synthesis (SIS) refers to the problem of generating reali...
research
03/13/2020

MixPoet: Diverse Poetry Generation via Learning Controllable Mixed Latent Space

As an essential step towards computer creativity, automatic poetry gener...

Please sign up or login with your details

Forgot password? Click here to reset