A Recurrent Encoder-Decoder Network for Sequential Face Alignment

08/19/2016
by   Xi Peng, et al.
0

We propose a novel recurrent encoder-decoder network model for real-time video-based face alignment. Our proposed model predicts 2D facial point maps regularized by a regression loss, while uniquely exploiting recurrent learning at both spatial and temporal dimensions. At the spatial level, we add a feedback loop connection between the combined output response map and the input, in order to enable iterative coarse-to-fine face alignment using a single network model. At the temporal level, we first decouple the features in the bottleneck of the network into temporal-variant factors, such as pose and expression, and temporal-invariant factors, such as identity information. Temporal recurrent learning is then applied to the decoupled temporal-variant features, yielding better generalization and significantly more accurate results at test time. We perform a comprehensive experimental analysis, showing the importance of each component of our proposed model, as well as superior results over the state-of-the-art in standard datasets.

READ FULL TEXT

page 3

page 7

page 9

research
01/17/2018

RED-Net: A Recurrent Encoder-Decoder Network for Video-based Face Alignment

We propose a novel method for real-time face alignment in videos based o...
research
09/19/2019

Dual Encoder-Decoder based Generative Adversarial Networks for Disentangled Facial Representation Learning

To learn disentangled representations of facial images, we present a Dua...
research
12/06/2016

Video Ladder Networks

We present the Video Ladder Network (VLN) for efficiently generating fut...
research
02/04/2022

Multi-task head pose estimation in-the-wild

We present a deep learning-based multi-task approach for head pose estim...
research
05/08/2019

Deep Blind Video Decaptioning by Temporal Aggregation and Recurrence

Blind video decaptioning is a problem of automatically removing text ove...
research
05/12/2017

Spatial-Temporal Recurrent Neural Network for Emotion Recognition

Emotion analysis is a crucial problem to endow artifact machines with re...
research
09/02/2015

What to talk about and how? Selective Generation using LSTMs with Coarse-to-Fine Alignment

We propose an end-to-end, domain-independent neural encoder-aligner-deco...

Please sign up or login with your details

Forgot password? Click here to reset