Semantic Flow for Fast and Accurate Scene Parsing

02/24/2020
by   Xiangtai Li, et al.
0

In this paper, we focus on effective methods for fast and accurate scene parsing. A common practice to improve the performance is to attain high resolution feature maps with strong semantic representation. Two strategies are widely used—astrous convolutions and feature pyramid fusion, are either computation intensive or ineffective. Inspired by Optical Flow for motion alignment between adjacent video frames, we propose a Flow Alignment Module (FAM) to learn Semantic Flow between feature maps of adjacent levels and broadcast high-level features to high resolution features effectively and efficiently. Furthermore, integrating our module to a common feature pyramid structure exhibits superior performance over other real-time methods even on very light-weight backbone networks, such as ResNet-18. Extensive experiments are conducted on several challenging datasets, including Cityscapes, PASCAL Context, ADE20K and CamVid. Particularly, our network is the first to achieve 80.4% mIoU on Cityscapes with a frame rate of 26 FPS. The code will be available at <https://github.com/donnyyou/torchcv>.

READ FULL TEXT

page 3

page 7

research
07/10/2022

SFNet: Faster, Accurate, and Domain Agnostic Semantic Segmentation via Semantic Flow

In this paper, we focus on exploring effective methods for faster, accur...
research
05/25/2021

Fast and Accurate Scene Parsing via Bi-direction Alignment Networks

In this paper, we propose an effective method for fast and accurate scen...
research
08/16/2021

FaPN: Feature-aligned Pyramid Network for Dense Image Prediction

Recent advancements in deep neural networks have made remarkable leap-fo...
research
03/28/2019

FastFCN: Rethinking Dilated Convolution in the Backbone for Semantic Segmentation

Modern approaches for semantic segmentation usually employ dilated convo...
research
12/09/2021

Edge-aware Guidance Fusion Network for RGB Thermal Scene Parsing

RGB thermal scene parsing has recently attracted increasing research int...
research
10/26/2020

P^2 Net: Augmented Parallel-Pyramid Net for Attention Guided Pose Estimation

We propose an augmented Parallel-Pyramid Net (P^2 Net) with feature refi...
research
05/29/2022

IFRNet: Intermediate Feature Refine Network for Efficient Frame Interpolation

Prevailing video frame interpolation algorithms, that generate the inter...

Please sign up or login with your details

Forgot password? Click here to reset