Paraphrase Generation as Zero-Shot Multilingual Translation:   Disentangling Semantic Similarity from Lexical and Syntactic Diversity

Brian Thompson; Matt Post

arXiv:2008.04935·cs.CL·October 29, 2020

Paraphrase Generation as Zero-Shot Multilingual Translation: Disentangling Semantic Similarity from Lexical and Syntactic Diversity

Brian Thompson, Matt Post

PDF

1 Repo

TL;DR

This paper presents a novel multilingual paraphrase generation method that controls lexical diversity and outperforms existing English paraphrasers in preserving meaning and grammaticality, using a single NMT model.

Contribution

It introduces a simple algorithm for multilingual paraphrase generation that discourages copying and allows lexical diversity control, improving quality over existing methods.

Findings

01

Produces paraphrases with better meaning preservation

02

Generates more grammatical paraphrases

03

Works effectively in multiple languages

Abstract

Recent work has shown that a multilingual neural machine translation (NMT) model can be used to judge how well a sentence paraphrases another sentence in the same language (Thompson and Post, 2020); however, attempting to generate paraphrases from such a model using standard beam search produces trivial copies or near copies. We introduce a simple paraphrase generation algorithm which discourages the production of n-grams that are present in the input. Our approach enables paraphrase generation in many languages from a single multilingual NMT model. Furthermore, the amount of lexical diversity between the input and output can be controlled at generation time. We conduct a human evaluation to compare our method to a paraphraser trained on the large English synthetic paraphrase database ParaBank 2 (Hu et al., 2019c) and find that our method produces paraphrases that better preserve…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

thompsonb/prism
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.