HINT3: Raising the bar for Intent Detection in the Wild

Gaurav Arora; Chirag Jain; Manas Chaturvedi; Krupal Modi

arXiv:2009.13833·cs.CL·March 25, 2021

HINT3: Raising the bar for Intent Detection in the Wild

Gaurav Arora, Chirag Jain, Manas Chaturvedi, Krupal Modi

PDF

1 Repo

TL;DR

This paper introduces three real-world chatbot datasets to improve intent detection benchmarking, revealing current systems' reliance on unintended patterns and highlighting the need for more robust models.

Contribution

The paper presents new real-user chatbot datasets and evaluates existing NLU systems, exposing their limitations in handling real-world intent detection complexities.

Findings

01

Current systems rely on unintended correlations in training data.

02

Performance saturates at low levels on real-world test sets.

03

Benchmarking with real datasets exposes robustness issues.

Abstract

Intent Detection systems in the real world are exposed to complexities of imbalanced datasets containing varying perception of intent, unintended correlations and domain-specific aberrations. To facilitate benchmarking which can reflect near real-world scenarios, we introduce 3 new datasets created from live chatbots in diverse domains. Unlike most existing datasets that are crowdsourced, our datasets contain real user queries received by the chatbots and facilitates penalising unwanted correlations grasped during the training process. We evaluate 4 NLU platforms and a BERT based classifier and find that performance saturates at inadequate levels on test sets because all systems latch on to unintended patterns in training data.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

hellohaptik/HINT3
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsLinear Layer · Softmax · Refunds@Expedia|||How do I get a full refund from Expedia? · Dense Connections · Dropout · Linear Warmup With Linear Decay · Layer Normalization · Attention Dropout · WordPiece · Weight Decay