Generating Multiple Objects at Spatially Distinct Locations

01/03/2019
by   Tobias Hinz, et al.
10

Recent improvements to Generative Adversarial Networks (GANs) have made it possible to generate realistic images in high resolution based on natural language descriptions such as image captions. Furthermore, conditional GANs allow us to control the image generation process through labels or even natural language descriptions. However, fine-grained control of the image layout, i.e. where in the image specific objects should be located, is still difficult to achieve. This is especially true for images that should contain multiple distinct objects at different spatial locations. We introduce a new approach which allows us to control the location of arbitrarily many objects within an image by adding an object pathway to both the generator and the discriminator. Our approach does not need a detailed semantic layout but only bounding boxes and the respective labels of the desired objects are needed. The object pathway focuses solely on the individual objects and is iteratively applied at the locations specified by the bounding boxes. The global pathway focuses on the image background and the general image layout. We perform experiments on the Multi-MNIST, CLEVR, and the more complex MS-COCO data set. Our experiments show that through the use of the object pathway we can control object locations within images and can model complex scenes with multiple objects at various locations. We further show that the object pathway focuses on the individual objects and learns features relevant for these, while the global pathway focuses on global image characteristics and the image background.

READ FULL TEXT

page 5

page 6

page 8

page 9

page 17

page 18

page 19

page 20

research
03/16/2020

Object-Centric Image Generation from Layouts

Despite recent impressive results on single-object and single-domain ima...
research
03/26/2020

BachGAN: High-Resolution Image Synthesis from Salient Object Layout

We propose a new task towards more practical application for image gener...
research
03/25/2021

AttrLostGAN: Attribute Controlled Image Synthesis from Reconfigurable Layout and Style

Conditional image synthesis from layout has recently attracted much inte...
research
04/26/2023

Controllable Image Generation via Collage Representations

Recent advances in conditional generative image models have enabled impr...
research
05/14/2017

GeneGAN: Learning Object Transfiguration and Attribute Subspace from Unpaired Data

Object Transfiguration replaces an object in an image with another objec...
research
02/04/2019

Realistic Image Generation using Region-phrase Attention

The Generative Adversarial Network (GAN) has recently been applied to ge...
research
08/20/2019

Image Synthesis From Reconfigurable Layout and Style

Despite remarkable recent progress on both unconditional and conditional...

Please sign up or login with your details

Forgot password? Click here to reset