Uncovering divergent linguistic information in word embeddings with   lessons for intrinsic and extrinsic evaluation

Mikel Artetxe; Gorka Labaka; I\~nigo Lopez-Gazpio; Eneko Agirre

arXiv:1809.02094·cs.CL·December 28, 2021

Uncovering divergent linguistic information in word embeddings with lessons for intrinsic and extrinsic evaluation

Mikel Artetxe, Gorka Labaka, I\~nigo Lopez-Gazpio, Eneko Agirre

PDF

2 Repos

TL;DR

This paper demonstrates that word embeddings encode more information than apparent, and that linear transformations can enhance their linguistic properties, revealing insights into intrinsic and extrinsic evaluation methods.

Contribution

It introduces a linear transformation approach to adjust similarity orders in embeddings, improving their performance in capturing linguistic aspects without external resources.

Findings

01

Transformations improve intrinsic linguistic property capture.

02

Transformations have a greater impact on unsupervised downstream tasks.

03

Embeddings encode more diverse information than initially apparent.

Abstract

Following the recent success of word embeddings, it has been argued that there is no such thing as an ideal representation for words, as different models tend to capture divergent and often mutually incompatible aspects like semantics/syntax and similarity/relatedness. In this paper, we show that each embedding model captures more information than directly apparent. A linear transformation that adjusts the similarity order of the model without any external resource can tailor it to achieve better results in those aspects, providing a new perspective on how embeddings encode divergent linguistic information. In addition, we explore the relation between intrinsic and extrinsic evaluation, as the effect of our transformations in downstream tasks is higher for unsupervised systems than for supervised ones.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.