(AF)2-S3Net: Attentive Feature Fusion with Adaptive Feature Selection for Sparse Semantic Segmentation Network

by   Ran Cheng, et al.

Autonomous robotic systems and self driving cars rely on accurate perception of their surroundings as the safety of the passengers and pedestrians is the top priority. Semantic segmentation is one the essential components of environmental perception that provides semantic information of the scene. Recently, several methods have been introduced for 3D LiDAR semantic segmentation. While, they can lead to improved performance, they are either afflicted by high computational complexity, therefore are inefficient, or lack fine details of smaller instances. To alleviate this problem, we propose AF2-S3Net, an end-to-end encoder-decoder CNN network for 3D LiDAR semantic segmentation. We present a novel multi-branch attentive feature fusion module in the encoder and a unique adaptive feature selection module with feature map re-weighting in the decoder. Our AF2-S3Net fuses the voxel based learning and point-based learning into a single framework to effectively process the large 3D scene. Our experimental results show that the proposed method outperforms the state-of-the-art approaches on the large-scale SemanticKITTI benchmark, ranking 1st on the competitive public leaderboard competition upon publication.


page 1

page 4

page 6

page 7


S3Net: 3D LiDAR Sparse Semantic Segmentation Network

Semantic Segmentation is a crucial component in the perception systems o...

Lite-HDSeg: LiDAR Semantic Segmentation Using Lite Harmonic Dense Convolutions

Autonomous driving vehicles and robotic systems rely on accurate percept...

LidarMultiNet: Unifying LiDAR Semantic Segmentation, 3D Object Detection, and Panoptic Segmentation in a Single Multi-task Network

This technical report presents the 1st place winning solution for the Wa...

VMNet: Voxel-Mesh Network for Geodesic-Aware 3D Semantic Segmentation

In recent years, sparse voxel-based methods have become the state-of-the...

Multimodal Sensor-Based Semantic 3D Mapping for a Large-Scale Environment

Semantic 3D mapping is one of the most important fields in robotics, and...

LidarMultiNet: Towards a Unified Multi-task Network for LiDAR Perception

LiDAR-based 3D object detection, semantic segmentation, and panoptic seg...

AMVNet: Assertion-based Multi-View Fusion Network for LiDAR Semantic Segmentation

In this paper, we present an Assertion-based Multi-View Fusion network (...