Getting Gender Right in Neural Machine Translation

Eva Vanmassenhove; Christian Hardmeier; Andy Way

arXiv:1909.05088·cs.CL·September 12, 2019

Getting Gender Right in Neural Machine Translation

Eva Vanmassenhove, Christian Hardmeier, Andy Way

PDF

TL;DR

This paper explores how incorporating gender information into neural machine translation improves translation accuracy, especially for languages with grammatical gender, by compiling speaker datasets and conducting experiments across 20 language pairs.

Contribution

It introduces large speaker datasets for 20 language pairs and demonstrates that adding gender features enhances NMT translation quality.

Findings

01

Gender feature integration improves translation accuracy for some language pairs.

02

Large datasets with speaker gender information are compiled for multiple languages.

03

Simple experiments show the benefit of gender-aware NMT models.

Abstract

Speakers of different languages must attend to and encode strikingly different aspects of the world in order to use their language correctly (Sapir, 1921; Slobin, 1996). One such difference is related to the way gender is expressed in a language. Saying "I am happy" in English, does not encode any additional knowledge of the speaker that uttered the sentence. However, many other languages do have grammatical gender systems and so such knowledge would be encoded. In order to correctly translate such a sentence into, say, French, the inherent gender information needs to be retained/recovered. The same sentence would become either "Je suis heureux", for a male speaker or "Je suis heureuse" for a female one. Apart from morphological agreement, demographic factors (gender, age, etc.) also influence our use of language in terms of word choices or even on the level of syntactic constructions…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsAttention Model