Efficient DETR: Improving End-to-End Object Detector with Dense Prior

04/03/2021
by   Zhuyu Yao, et al.
0

The recently proposed end-to-end transformer detectors, such as DETR and Deformable DETR, have a cascade structure of stacking 6 decoder layers to update object queries iteratively, without which their performance degrades seriously. In this paper, we investigate that the random initialization of object containers, which include object queries and reference points, is mainly responsible for the requirement of multiple iterations. Based on our findings, we propose Efficient DETR, a simple and efficient pipeline for end-to-end object detection. By taking advantage of both dense detection and sparse set detection, Efficient DETR leverages dense prior to initialize the object containers and brings the gap of the 1-decoder structure and 6-decoder structure. Experiments conducted on MS COCO show that our method, with only 3 encoder layers and 1 decoder layer, achieves competitive performance with state-of-the-art object detection methods. Efficient DETR is also robust in crowded scenes. It outperforms modern detectors on CrowdHuman dataset by a large margin.

READ FULL TEXT

page 3

page 4

page 6

research
03/22/2023

Dense Distinct Query for End-to-End Object Detection

One-to-one label assignment in object detection has successfully obviate...
research
06/02/2022

What Are Expected Queries in End-to-End Object Detection?

End-to-end object detection is rapidly progressed after the emergence of...
research
05/24/2019

A Real-Time Tiny Detection Model for Stem End and Blossom End of Navel Orange

To distinguish the stem end and blossom end of navel orange from its bla...
research
04/17/2023

DETRs Beat YOLOs on Real-time Object Detection

Recently, end-to-end transformer-based detectors (DETRs) have achieved r...
research
11/29/2021

Sparse DETR: Efficient End-to-End Object Detection with Learnable Sparsity

DETR is the first end-to-end object detector using a transformer encoder...
research
03/01/2023

D2Q-DETR: Decoupling and Dynamic Queries for Oriented Object Detection with Transformers

Despite the promising results, existing oriented object detection method...
research
05/23/2021

COTR: Convolution in Transformer Network for End to End Polyp Detection

Purpose: Colorectal cancer (CRC) is the second most common cause of canc...

Please sign up or login with your details

Forgot password? Click here to reset