A Geometric Analysis of Neural Collapse with Unconstrained Features

Zhihui Zhu; Tianyu Ding; Jinxin Zhou; Xiao Li; Chong You; Jeremias; Sulam; and Qing Qu

arXiv:2105.02375·cs.LG·May 7, 2021·43 cites

A Geometric Analysis of Neural Collapse with Unconstrained Features

Zhihui Zhu, Tianyu Ding, Jinxin Zhou, Xiao Li, Chong You, Jeremias, Sulam, and Qing Qu

PDF

Open Access 1 Repo 1 Video

TL;DR

This paper analyzes the global optimization landscape of Neural Collapse in neural networks, showing that simple models with cross-entropy loss and weight decay naturally lead to class features forming a Simplex ETF, explaining empirical phenomena and enabling efficient training.

Contribution

It provides the first global landscape analysis of Neural Collapse using a simplified unconstrained feature model, linking theoretical insights to practical neural network training.

Findings

01

Global minimizers are Simplex ETFs

02

Critical points are strict saddles with negative curvature

03

Fixing features as Simplex ETF reduces memory without losing performance

Abstract

We provide the first global optimization landscape analysis of $N e u r a l C o l l a p se$ -- an intriguing empirical phenomenon that arises in the last-layer classifiers and features of neural networks during the terminal phase of training. As recently reported by Papyan et al., this phenomenon implies that ( $i$ ) the class means and the last-layer classifiers all collapse to the vertices of a Simplex Equiangular Tight Frame (ETF) up to scaling, and ( $ii$ ) cross-example within-class variability of last-layer activations collapses to zero. We study the problem based on a simplified $u n co n s t r ain e d f e a t u r e m o d e l$ , which isolates the topmost layers from the classifier of the neural network. In this context, we show that the classical cross-entropy loss with weight decay has a benign global landscape, in the sense that the only global minimizers are the Simplex ETFs while all other critical points…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

tding1/Neural-Collapse
pytorchOfficial

Videos

A Geometric Analysis of Neural Collapse with Unconstrained Features· slideslive

Taxonomy

TopicsMedical Image Segmentation Techniques · Cell Image Analysis Techniques · Morphological variations and asymmetry

MethodsWeight Decay