Text Counterfactuals via Latent Optimization and Shapley-Guided Search

Quintin Pope; Xiaoli Z. Fern

arXiv:2110.11589·cs.CL·October 25, 2021

Text Counterfactuals via Latent Optimization and Shapley-Guided Search

Quintin Pope, Xiaoli Z. Fern

PDF

Open Access 1 Repo

TL;DR

This paper introduces a novel method for generating counterfactual texts by optimizing in latent space and using Shapley values to guide modifications, improving interpretability and debugging of classifiers.

Contribution

It proposes a new approach combining latent optimization and Shapley-guided search for counterfactual text generation, addressing challenges of discrete text modifications.

Findings

01

Outperforms recent baselines in success rate and quality

02

Latent optimization improves counterfactual relevance

03

Shapley values enhance the effectiveness of modifications

Abstract

We study the problem of generating counterfactual text for a classifier as a means for understanding and debugging classification. Given a textual input and a classification model, we aim to minimally alter the text to change the model's prediction. White-box approaches have been successfully applied to similar problems in vision where one can directly optimize the continuous input. Optimization-based approaches become difficult in the language domain due to the discrete nature of text. We bypass this issue by directly optimizing in the latent space and leveraging a language model to generate candidate modifications from optimized latent representations. We additionally use Shapley values to estimate the combinatoric effect of multiple changes. We then use these estimates to guide a beam search for the final counterfactual text. We achieve favorable performance compared to recent…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

QuintinPope/CLOSS
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsTopic Modeling · Multimodal Machine Learning Applications · Domain Adaptation and Few-Shot Learning