Tighter Expected Generalization Error Bounds via Convexity of   Information Measures

Gholamali Aminian; Yuheng Bu; Gregory Wornell; Miguel Rodrigues

arXiv:2202.12150·cs.IT·February 25, 2022·1 cites

Tighter Expected Generalization Error Bounds via Convexity of Information Measures

Gholamali Aminian, Yuheng Bu, Gregory Wornell, Miguel Rodrigues

PDF

Open Access

TL;DR

This paper introduces new expected generalization error bounds using convex information measures, resulting in tighter bounds than previous methods, with applications demonstrated through examples.

Contribution

It develops novel generalization bounds based on convex information measures like Wasserstein and total variation distances, improving tightness over existing bounds.

Findings

01

Bounds based on Wasserstein and total variation are tighter.

02

Convexity of information measures enhances bound tightness.

03

Example demonstrates the bounds' effectiveness.

Abstract

Generalization error bounds are essential to understanding machine learning algorithms. This paper presents novel expected generalization error upper bounds based on the average joint distribution between the output hypothesis and each input training sample. Multiple generalization error upper bounds based on different information measures are provided, including Wasserstein distance, total variation distance, KL divergence, and Jensen-Shannon divergence. Due to the convexity of the information measures, the proposed bounds in terms of Wasserstein distance and total variation distance are shown to be tighter than their counterparts based on individual samples in the literature. An example is provided to demonstrate the tightness of the proposed generalization error bounds.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsSparse and Compressive Sensing Techniques · Face and Expression Recognition · Domain Adaptation and Few-Shot Learning