Frame-To-Frame Consistent Semantic Segmentation

08/03/2020
by   Manuel Rebol, et al.
8

In this work, we aim for temporally consistent semantic segmentation throughout frames in a video. Many semantic segmentation algorithms process images individually which leads to an inconsistent scene interpretation due to illumination changes, occlusions and other variations over time. To achieve a temporally consistent prediction, we train a convolutional neural network (CNN) which propagates features through consecutive frames in a video using a convolutional long short term memory (ConvLSTM) cell. Besides the temporal feature propagation, we penalize inconsistencies in our loss function. We show in our experiments that the performance improves when utilizing video information compared to single frame prediction. The mean intersection over union (mIoU) metric on the Cityscapes validation set increases from 45.2 the single frames to 57.9 propagate features trough time on the ESPNet. Most importantly, inconsistency decreases from 4.5 indicate that the added temporal information produces a frame-to-frame consistent and more accurate image understanding compared to single frame processing.

READ FULL TEXT

page 1

page 4

page 6

research
01/05/2021

Local Memory Attention for Fast Video Semantic Segmentation

We propose a novel neural network module that transforms an existing sin...
research
11/20/2020

Recovering the Imperfect: Cell Segmentation in the Presence of Dynamically Localized Proteins

Deploying off-the-shelf segmentation networks on biomedical data has bec...
research
05/24/2019

Implicit Label Augmentation on Partially Annotated Clips via Temporally-Adaptive Features Learning

Partially annotated clips contain rich temporal contexts that can comple...
research
03/03/2021

LiDAR-based Recurrent 3D Semantic Segmentation with Temporal Memory Alignment

Understanding and interpreting a 3d environment is a key challenge for a...
research
10/24/2021

Perceptual Consistency in Video Segmentation

In this paper, we present a novel perceptual consistency perspective on ...
research
04/03/2020

Temporally Distributed Networks for Fast Video Segmentation

We present TDNet, a temporally distributed network designed for fast and...
research
04/03/2020

Temporally Distributed Networks for Fast Video Semantic Segmentation

We present TDNet, a temporally distributed network designed for fast and...

Please sign up or login with your details

Forgot password? Click here to reset