Leveraging in-domain supervision for unsupervised image-to-image translation tasks via multi-stream generators

12/30/2021
by   Dvir Yerushalmi, et al.
0

Supervision for image-to-image translation (I2I) tasks is hard to come by, but bears significant effect on the resulting quality. In this paper, we observe that for many Unsupervised I2I (UI2I) scenarios, one domain is more familiar than the other, and offers in-domain prior knowledge, such as semantic segmentation. We argue that for complex scenes, figuring out the semantic structure of the domain is hard, especially with no supervision, but is an important part of a successful I2I operation. We hence introduce two techniques to incorporate this invaluable in-domain prior knowledge for the benefit of translation quality: through a novel Multi-Stream generator architecture, and through a semantic segmentation-based regularization loss term. In essence, we propose splitting the input data according to semantic masks, explicitly guiding the network to different behavior for the different regions of the image. In addition, we propose training a semantic segmentation network along with the translation task, and to leverage this output as a loss term that improves robustness. We validate our approach on urban data, demonstrating superior quality in the challenging UI2I tasks of converting day images to night ones. In addition, we also demonstrate how reinforcing the target dataset with our augmented images improves the training of downstream tasks such as the classical detection one.

READ FULL TEXT

page 3

page 4

page 5

page 6

page 7

page 8

page 9

page 11

research
10/11/2018

SingleGAN: Image-to-Image Translation by a Single-Generator Network using Multiple Generative Adversarial Learning

Image translation is a burgeoning field in computer vision where the goa...
research
02/11/2019

Unpaired Image-to-Image Translation with Domain Supervision

Image-to-image translation has been widely investigated in recent years....
research
09/23/2022

Image-to-Image Translation for Autonomous Driving from Coarsely-Aligned Image Pairs

A self-driving car must be able to reliably handle adverse weather condi...
research
09/06/2023

Exploring Semantic Consistency in Unpaired Image Translation to Generate Data for Surgical Applications

In surgical computer vision applications, obtaining labeled training dat...
research
08/13/2018

Improving Shape Deformation in Unsupervised Image-to-Image Translation

Unsupervised image-to-image translation techniques are able to map local...
research
12/09/2020

Lipschitz Regularized CycleGAN for Improving Semantic Robustness in Unpaired Image-to-image Translation

For unpaired image-to-image translation tasks, GAN-based approaches are ...
research
02/13/2023

DEPAS: De-novo Pathology Semantic Masks using a Generative Model

The integration of artificial intelligence into digital pathology has th...

Please sign up or login with your details

Forgot password? Click here to reset