Bayesian Topic Regression for Causal Inference

Maximilian Ahrens; Julian Ashwin; Jan-Peter Calliess; Vu Nguyen

arXiv:2109.05317·stat.ML·September 14, 2021

Bayesian Topic Regression for Causal Inference

Maximilian Ahrens, Julian Ashwin, Jan-Peter Calliess, Vu Nguyen

PDF

1 Repo

TL;DR

This paper introduces Bayesian Topic Regression (BTR), a novel model combining text and numerical data for causal inference, improving bias reduction and prediction accuracy over benchmarks.

Contribution

It develops a joint Bayesian framework for causal inference with text and numerical confounders, outperforming existing methods in bias reduction and prediction.

Findings

01

Lower bias in ground truth recovery with synthetic data

02

Superior prediction accuracy on real-world datasets

03

Competitive with complex deep neural networks

Abstract

Causal inference using observational text data is becoming increasingly popular in many research areas. This paper presents the Bayesian Topic Regression (BTR) model that uses both text and numerical information to model an outcome variable. It allows estimation of both discrete and continuous treatment effects. Furthermore, it allows for the inclusion of additional numerical confounding factors next to text data. To this end, we combine a supervised Bayesian topic model with a Bayesian regression framework and perform supervised representation learning for the text features jointly with the regression parameter training, respecting the Frisch-Waugh-Lovell theorem. Our paper makes two main contributions. First, we provide a regression framework that allows causal inference in settings when both text and numerical confounders are of relevance. We show with synthetic and semi-synthetic…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

maximilianahrens/data
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.