Neural Machine Translation Doesn't Translate Gender Coreference Right Unless You Make It

10/11/2020
by   Danielle Saunders, et al.
0

Neural Machine Translation (NMT) has been shown to struggle with grammatical gender that is dependent on the gender of human referents, which can cause gender bias effects. Many existing approaches to this problem seek to control gender inflection in the target language by explicitly or implicitly adding a gender feature to the source sentence, usually at the sentence level. In this paper we propose schemes for incorporating explicit word-level gender inflection tags into NMT. We explore the potential of this gender-inflection controlled translation when the gender feature can be determined from a human reference, assessing on English-to-Spanish and English-to-German translation. We find that simple existing approaches can over-generalize a gender-feature to multiple entities in a sentence, and suggest an effective alternative in the form of tagged coreference adaptation data. We also propose an extension to assess translations of gender-neutral entities from English given a corresponding linguistic convention in the inflected target language.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
09/11/2019

Getting Gender Right in Neural Machine Translation

Speakers of different languages must attend to and encode strikingly dif...
research
02/26/2018

Gender Aware Spoken Language Translation Applied to English-Arabic

Spoken Language Translation (SLT) is becoming more widely used and becom...
research
08/05/2021

GENder-IT: An Annotated English-Italian Parallel Challenge Set for Cross-Linguistic Natural Gender Phenomena

Languages differ in terms of the absence or presence of gender features,...
research
04/15/2021

First the worst: Finding better gender translations during beam search

Neural machine translation inference procedures like beam search generat...
research
04/16/2021

Investigating Failures of Automatic Translation in the Case of Unambiguous Gender

Transformer based models are the modern work horses for neural machine t...
research
07/13/2021

Generating Gender Augmented Data for NLP

Gender bias is a frequent occurrence in NLP-based applications, especial...
research
10/13/2020

Mitigating Gender Bias in Machine Translation with Target Gender Annotations

When translating "The secretary asked for details." to a language with g...

Please sign up or login with your details

Forgot password? Click here to reset