Domain Feature Collapse: Implications for Out-of-Distribution Detection and Solutions

Hong Yang; Devroop Kar; Qi Yu; Alex Ororbia; Travis Desell

arXiv:2512.04034·cs.LG·March 13, 2026

Domain Feature Collapse: Implications for Out-of-Distribution Detection and Solutions

Hong Yang, Devroop Kar, Qi Yu, Alex Ororbia, Travis Desell

PDF

Open Access

TL;DR

This paper provides a theoretical explanation for why out-of-distribution detection fails in single-domain models, showing that models discard domain-specific features due to information bottleneck effects, and proposes domain filtering as a solution.

Contribution

It introduces the concept of domain feature collapse caused by information bottleneck in supervised learning, supported by a new benchmark and empirical validation.

Findings

01

Supervised learning on single domains leads to discarding domain features.

02

Preserving domain information improves out-of-distribution detection.

03

Domain filtering with pretrained representations mitigates failure modes.

Abstract

Why do state-of-the-art OOD detection methods exhibit catastrophic failure when models are trained on single-domain datasets? We provide the first theoretical explanation for this phenomenon through the lens of information theory. We prove that supervised learning on single-domain data inevitably produces domain feature collapse -- representations where I(x_d; z) = 0, meaning domain-specific information is completely discarded. This is a fundamental consequence of information bottleneck optimization: models trained on single domains (e.g., medical images) learn to rely solely on class-specific features while discarding domain features, leading to catastrophic failure when detecting out-of-domain samples (e.g., achieving only 53% FPR@95 on MNIST). We extend our analysis using Fano's inequality to quantify partial collapse in practical scenarios. To validate our theory, we introduce…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsDomain Adaptation and Few-Shot Learning · Advanced Neural Network Applications · Multimodal Machine Learning Applications