When and Why does a Model Fail? A Human-in-the-loop Error Detection   Framework for Sentiment Analysis

Zhe Liu; Yufan Guo; Jalal Mahmud

arXiv:2106.00954·cs.CL·June 3, 2021·5 cites

When and Why does a Model Fail? A Human-in-the-loop Error Detection Framework for Sentiment Analysis

Zhe Liu, Yufan Guo, Jalal Mahmud

PDF

Open Access

TL;DR

This paper introduces a human-in-the-loop framework for detecting errors in sentiment analysis models using explainable features, improving error identification before and after deployment.

Contribution

It presents a novel error detection framework combining global and local feature analysis with human assessment, enhancing model reliability.

Findings

01

High precision in identifying erroneous predictions with limited human intervention

02

Effective global and local feature contribution analysis for error detection

03

Improved model assessment prior to deployment

Abstract

Although deep neural networks have been widely employed and proven effective in sentiment analysis tasks, it remains challenging for model developers to assess their models for erroneous predictions that might exist prior to deployment. Once deployed, emergent errors can be hard to identify in prediction run-time and impossible to trace back to their sources. To address such gaps, in this paper we propose an error detection framework for sentiment analysis based on explainable features. We perform global-level feature validation with human-in-the-loop assessment, followed by an integration of global and local-level feature contribution analysis. Experimental results show that, given limited human-in-the-loop intervention, our method is able to identify erroneous model predictions on unseen data with high precision.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdversarial Robustness in Machine Learning · Anomaly Detection Techniques and Applications · Explainable Artificial Intelligence (XAI)