ConjNLI: Natural Language Inference Over Conjunctive Sentences

Swarnadeep Saha; Yixin Nie; Mohit Bansal

arXiv:2010.10418·cs.CL·October 23, 2020

ConjNLI: Natural Language Inference Over Conjunctive Sentences

Swarnadeep Saha, Yixin Nie, Mohit Bansal

PDF

1 Repo

TL;DR

ConjNLI introduces a challenging stress-test for natural language inference involving conjunctive sentences, revealing that current models like RoBERTa struggle with conjunctive semantics and require further development.

Contribution

The paper presents ConjNLI, a novel benchmark for testing NLI over conjunctive sentences, and proposes adversarial fine-tuning and predicate role awareness to improve model understanding.

Findings

01

Large models like RoBERTa rely on shallow heuristics for conjunctive inference.

02

Adversarial fine-tuning improves model performance on ConjNLI.

03

ConjNLI remains challenging, indicating room for further research.

Abstract

Reasoning about conjuncts in conjunctive sentences is important for a deeper understanding of conjunctions in English and also how their usages and semantics differ from conjunctive and disjunctive boolean logic. Existing NLI stress tests do not consider non-boolean usages of conjunctions and use templates for testing such model knowledge. Hence, we introduce ConjNLI, a challenge stress-test for natural language inference over conjunctive sentences, where the premise differs from the hypothesis by conjuncts removed, added, or replaced. These sentences contain single and multiple instances of coordinating conjunctions ("and", "or", "but", "nor") with quantifiers, negations, and requiring diverse boolean and non-boolean inferences over conjuncts. We find that large-scale pre-trained language models like RoBERTa do not understand conjunctive semantics well and resort to shallow heuristics…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

swarnaHub/ConjNLI
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsLinear Layer · Adam · Layer Normalization · Dense Connections · Multi-Head Attention · Refunds@Expedia|||How do I get a full refund from Expedia? · Dropout · Linear Warmup With Linear Decay · Attention Dropout · Weight Decay