Improving Human Image Synthesis with Residual Fast Fourier Transformation and Wasserstein Distance

05/24/2022
by   Jianhan Wu, et al.
0

With the rapid development of the Metaverse, virtual humans have emerged, and human image synthesis and editing techniques, such as pose transfer, have recently become popular. Most of the existing techniques rely on GANs, which can generate good human images even with large variants and occlusions. But from our best knowledge, the existing state-of-the-art method still has the following problems: the first is that the rendering effect of the synthetic image is not realistic, such as poor rendering of some regions. And the second is that the training of GAN is unstable and slow to converge, such as model collapse. Based on the above two problems, we propose several methods to solve them. To improve the rendering effect, we use the Residual Fast Fourier Transform Block to replace the traditional Residual Block. Then, spectral normalization and Wasserstein distance are used to improve the speed and stability of GAN training. Experiments demonstrate that the methods we offer are effective at solving the problems listed above, and we get state-of-the-art scores in LPIPS and PSNR.

READ FULL TEXT

page 2

page 6

page 7

research
03/06/2021

PISE: Person Image Synthesis and Editing with Decoupled GAN

Person image synthesis, e.g., pose transfer, is a challenging problem du...
research
05/02/2018

Text to Image Synthesis Using Generative Adversarial Networks

Generating images from natural language is one of the primary applicatio...
research
11/23/2021

Deep Residual Fourier Transformation for Single Image Deblurring

It has been a common practice to adopt the ResBlock, which learns the di...
research
06/18/2018

Banach Wasserstein GAN

Wasserstein Generative Adversarial Networks (WGANs) can be used to gener...
research
03/06/2019

Conditional GANs For Painting Generation

We examined the use of modern Generative Adversarial Nets to generate no...
research
01/05/2020

A Robust Pose Transformational GAN for Pose Guided Person Image Synthesis

Generating photorealistic images of human subjects in any unseen pose ha...

Please sign up or login with your details

Forgot password? Click here to reset