Overcoming the Rare Word Problem for Low-Resource Language Pairs in   Neural Machine Translation

Thi-Vinh Ngo; Thanh-Le Ha; Phuong-Thai Nguyen; Le-Minh Nguyen

arXiv:1910.03467·cs.CL·December 17, 2020

Overcoming the Rare Word Problem for Low-Resource Language Pairs in Neural Machine Translation

Thi-Vinh Ngo, Thanh-Le Ha, Phuong-Thai Nguyen, Le-Minh Nguyen

PDF

TL;DR

This paper introduces three novel methods to mitigate the rare word problem in neural machine translation for low-resource languages, improving translation quality by up to 1 BLEU point.

Contribution

The paper presents three innovative solutions: enhancing source context, learning morphology of unknown words, and leveraging WordNet synonyms to address rare words in NMT.

Findings

01

Achieved up to +1.0 BLEU point improvement

02

Effective in low-resource English-Vietnamese and Japanese-Vietnamese translation

03

Demonstrated significant reduction in rare word errors

Abstract

Among the six challenges of neural machine translation (NMT) coined by (Koehn and Knowles, 2017), rare-word problem is considered the most severe one, especially in translation of low-resource languages. In this paper, we propose three solutions to address the rare words in neural machine translation systems. First, we enhance source context to predict the target words by connecting directly the source embeddings to the output of the attention component in NMT. Second, we propose an algorithm to learn morphology of unknown words for English in supervised way in order to minimize the adverse effect of rare-word problem. Finally, we exploit synonymous relation from the WordNet to overcome out-of-vocabulary (OOV) problem of NMT. We evaluate our approaches on two low-resource language pairs: English-Vietnamese and Japanese-Vietnamese. In our experiments, we have achieved significant…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.