A comprehensive, application-oriented study of catastrophic forgetting   in DNNs

B. Pf\"ulb; A. Gepperth

arXiv:1905.08101·cs.LG·September 11, 2019·58 cites

A comprehensive, application-oriented study of catastrophic forgetting in DNNs

B. Pf\"ulb, A. Gepperth

PDF

Open Access 1 Repo

TL;DR

This paper conducts a large-scale empirical study on catastrophic forgetting in deep neural networks during sequential learning, revealing that no existing model completely avoids CF across diverse datasets and tasks.

Contribution

It introduces a new experimental protocol for realistic application scenarios and evaluates CF across the largest set of visual datasets to date.

Findings

01

No model completely avoids CF across all datasets and tasks

02

EWC and IMM models have potential workarounds for CF

03

Empirical evidence highlights the challenge of mitigating CF in practical settings

Abstract

We present a large-scale empirical study of catastrophic forgetting (CF) in modern Deep Neural Network (DNN) models that perform sequential (or: incremental) learning. A new experimental protocol is proposed that enforces typical constraints encountered in application scenarios. As the investigation is empirical, we evaluate CF behavior on the hitherto largest number of visual classification datasets, from each of which we construct a representative number of Sequential Learning Tasks (SLTs) in close alignment to previous works on CF. Our results clearly indicate that there is no model that avoids CF for all investigated datasets and SLTs under application conditions. We conclude with a discussion of potential solutions and workarounds to CF, notably for the EWC and IMM models.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

BPfuelb/CF_in_DNNs
tf

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsDomain Adaptation and Few-Shot Learning · Multimodal Machine Learning Applications

MethodsElastic Weight Consolidation