SeedBERT: Recovering Annotator Rating Distributions from an Aggregated   Label

Aneesha Sampath; Victoria Lin; Louis-Philippe Morency

arXiv:2211.13196·cs.LG·November 24, 2022·1 cites

SeedBERT: Recovering Annotator Rating Distributions from an Aggregated Label

Aneesha Sampath, Victoria Lin, Louis-Philippe Morency

PDF

Open Access

TL;DR

SeedBERT is a novel method that recovers annotator rating distributions from single labels by inducing pre-trained models to attend to different input parts, improving performance on subjective tasks.

Contribution

It introduces SeedBERT, a technique that infers annotator disagreement distributions from one label, addressing the lack of annotator-specific data in subjective datasets.

Findings

01

SeedBERT's attention aligns with human annotator disagreement.

02

It outperforms standard models on subjective tasks.

03

Demonstrates significant performance gains in empirical evaluations.

Abstract

Many machine learning tasks -- particularly those in affective computing -- are inherently subjective. When asked to classify facial expressions or to rate an individual's attractiveness, humans may disagree with one another, and no single answer may be objectively correct. However, machine learning datasets commonly have just one "ground truth" label for each sample, so models trained on these labels may not perform well on tasks that are subjective in nature. Though allowing models to learn from the individual annotators' ratings may help, most datasets do not provide annotator-specific labels for each sample. To address this issue, we propose SeedBERT, a method for recovering annotator rating distributions from a single label by inducing pre-trained models to attend to different portions of the input. Our human evaluations indicate that SeedBERT's attention mechanism is consistent…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsMisinformation and Its Impacts · Sentiment Analysis and Opinion Mining · Emotion and Mood Recognition