Investigate the Essence of Long-Tailed Recognition from a Unified Perspective

07/08/2021
by   Lei Liu, et al.
0

As the data scale grows, deep recognition models often suffer from long-tailed data distributions due to the heavy imbalanced sample number across categories. Indeed, real-world data usually exhibit some similarity relation among different categories (e.g., pigeons and sparrows), called category similarity in this work. It is doubly difficult when the imbalance occurs between such categories with similar appearances. However, existing solutions mainly focus on the sample number to re-balance data distribution. In this work, we systematically investigate the essence of the long-tailed problem from a unified perspective. Specifically, we demonstrate that long-tailed recognition suffers from both sample number and category similarity. Intuitively, using a toy example, we first show that sample number is not the unique influence factor for performance dropping of long-tailed recognition. Theoretically, we demonstrate that (1) category similarity, as an inevitable factor, would also influence the model learning under long-tailed distribution via similar samples, (2) using more discriminative representation methods (e.g., self-supervised learning) for similarity reduction, the classifier bias can be further alleviated with greatly improved performance. Extensive experiments on several long-tailed datasets verify the rationality of our theoretical analysis, and show that based on existing state-of-the-arts (SOTAs), the performance could be further improved by similarity reduction. Our investigations highlight the essence behind the long-tailed problem, and claim several feasible directions for future work.

READ FULL TEXT

page 5

page 7

research
06/13/2020

Rethinking the Value of Labels for Improving Class-Imbalanced Learning

Real-world data often exhibits long-tailed distributions with heavy clas...
research
12/15/2021

Imagine by Reasoning: A Reasoning-Based Implicit Semantic Data Augmentation for Long-Tailed Classification

Real-world data often follows a long-tailed distribution, which makes th...
research
11/06/2021

Towards Calibrated Model for Long-Tailed Visual Recognition from Prior Perspective

Real-world data universally confronts a severe class-imbalance problem a...
research
05/27/2022

A Survey on Long-Tailed Visual Recognition

The heavy reliance on data is one of the major reasons that currently li...
research
07/26/2022

Class-Aware Universum Inspired Re-Balance Learning for Long-Tailed Recognition

Data augmentation for minority classes is an effective strategy for long...
research
03/29/2022

Nested Collaborative Learning for Long-Tailed Visual Recognition

The networks trained on the long-tailed dataset vary remarkably, despite...
research
08/07/2022

Sample hardness based gradient loss for long-tailed cervical cell detection

Due to the difficulty of cancer samples collection and annotation, cervi...

Please sign up or login with your details

Forgot password? Click here to reset