Parameter sharing between dependency parsers for related languages

Miryam de Lhoneux; Johannes Bjerva; Isabelle Augenstein; Anders; S{\o}gaard

arXiv:1808.09055·cs.CL·October 8, 2018

Parameter sharing between dependency parsers for related languages

Miryam de Lhoneux, Johannes Bjerva, Isabelle Augenstein, Anders, S{\o}gaard

PDF

Open Access 1 Repo

TL;DR

This paper evaluates various parameter sharing strategies in neural dependency parsers for related languages, proposing a linguistically motivated architecture that improves performance over monolingual models.

Contribution

It systematically compares 27 sharing strategies across multiple language pairs and introduces a tunable sharing architecture for better parser performance.

Findings

01

Sharing transition classifier parameters consistently improves performance.

02

Selective sharing of word and character LSTM parameters varies in usefulness.

03

The proposed architecture outperforms monolingual baselines and adapts to related and unrelated languages.

Abstract

Previous work has suggested that parameter sharing between transition-based neural dependency parsers for related languages can lead to better performance, but there is no consensus on what parameters to share. We present an evaluation of 27 different parameter sharing strategies across 10 languages, representing five pairs of related languages, each pair from a different language family. We find that sharing transition classifier parameters always helps, whereas the usefulness of sharing word and/or character LSTM parameters varies. Based on this result, we propose an architecture where the transition classifier is shared, and the sharing of word and character parameters is controlled by a parameter that can be tuned on validation data. This model is linguistically motivated and obtains significant improvements over a monolingually trained baseline. We also find that sharing transition…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

coastalcph/uuparser
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsTopic Modeling · Natural Language Processing Techniques · Domain Adaptation and Few-Shot Learning

MethodsSigmoid Activation · Tanh Activation · Long Short-Term Memory