Scale-invariant representation of machine learning

Sungyeop Lee; Junghyo Jo

arXiv:2109.02914·cs.LG·April 13, 2022

Scale-invariant representation of machine learning

Sungyeop Lee, Junghyo Jo

PDF

1 Repo

TL;DR

This paper explores how machine learning models naturally develop scale-invariant internal representations characterized by power-law distributions, which reflect data compression and differentiation of outliers, rooted in information theory.

Contribution

It derives the theoretical process explaining the emergence of power-law distributions in machine learning representations, linking data compression and uncertainty.

Findings

01

Internal codes follow power-law distributions in models.

02

Scale-invariance relates to maximally uncertain data groupings.

03

Representation efficiency balances compression and differentiation.

Abstract

The success of machine learning has resulted from its structured representation of data. Similar data have close internal representations as compressed codes for classification or emerged labels for clustering. We observe that the frequency of internal codes or labels follows power laws in both supervised and unsupervised learning models. This scale-invariant distribution implies that machine learning largely compresses frequent typical data, and simultaneously, differentiates many atypical data as outliers. In this study, we derive the process by which these power laws can naturally arise in machine learning. In terms of information theory, the scale-invariant representation corresponds to a maximally uncertain data grouping among possible representations that guarantee a given learning accuracy.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

sungyeop/powerlaw_ml
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.