Do Neural Ranking Models Intensify Gender Bias?

05/01/2020
by   Navid Rekabsaz, et al.
0

Concerns regarding the footprint of societal biases in information retrieval (IR) systems have been raised in several previous studies. In this work, we examine various recent IR models from the perspective of the degree of gender bias in their retrieval results. To this end, we first provide a bias measurement framework which includes two metrics to quantify the degree of the unbalanced presence of gender-related concepts in a given IR model's ranking list. To examine IR models by means of the framework, we create a dataset of non-gendered queries, selected by human annotators. Applying these queries to the MS MARCO Passage retrieval collection, we then measure the gender bias of a BM25 model and several recent neural ranking models. The results show that while all models are strongly biased toward male, the neural models, and in particular the ones based on contextualized embedding models, significantly intensify gender bias. Our experiments also show an overall increase in the gender bias of neural models when they exploit transfer learning, namely when they use (already biased) pre-trained embeddings.

READ FULL TEXT
research
01/19/2022

Grep-BiasIR: A Dataset for Investigating Gender Representation-Bias in Information Retrieval Results

The provided contents by information retrieval (IR) systems can reflect ...
research
09/02/2020

Gender Stereotype Reinforcement: Measuring the Gender Bias Conveyed by Ranking Algorithms

Search Engines (SE) have been shown to perpetuate well-known gender ster...
research
05/18/2022

Debiasing Neural Retrieval via In-batch Balancing Regularization

People frequently interact with information retrieval (IR) systems, howe...
research
09/13/2019

Recommendation or Discrimination?: Quantifying Distribution Parity in Information Retrieval Systems

Information retrieval (IR) systems often leverage query data to suggest ...
research
01/21/2021

Assessing the Benefits of Model Ensembles in Neural Re-Ranking for Passage Retrieval

Our work aimed at experimentally assessing the benefits of model ensembl...
research
04/28/2021

Societal Biases in Retrieved Contents: Measurement Framework and Adversarial Mitigation for BERT Rankers

Societal biases resonate in the retrieved contents of information retrie...
research
04/29/2019

On the Effect of Low-Frequency Terms on Neural-IR Models

Low-frequency terms are a recurring challenge for information retrieval ...

Please sign up or login with your details

Forgot password? Click here to reset