Stochastic Weight Matrix-based Regularization Methods for Deep Neural   Networks

Patrik Reizinger; B\'alint Gyires-T\'oth

arXiv:1909.11977·cs.LG·June 7, 2022

Stochastic Weight Matrix-based Regularization Methods for Deep Neural Networks

Patrik Reizinger, B\'alint Gyires-T\'oth

PDF

1 Repo

TL;DR

This paper introduces two novel regularization techniques for deep neural networks based on weight matrix modifications, demonstrating improved performance and entropy on various datasets and tasks.

Contribution

The paper proposes two new regularization methods, Weight Reinitialization and Weight Shuffling, based on weight matrix modifications, with demonstrated effectiveness across multiple datasets.

Findings

01

Improved performance on MNIST, CIFAR-10, and JSB Chorales datasets.

02

Enhanced entropy and robustness of neural networks.

03

Methods applicable to time series modeling tasks.

Abstract

The aim of this paper is to introduce two widely applicable regularization methods based on the direct modification of weight matrices. The first method, Weight Reinitialization, utilizes a simplified Bayesian assumption with partially resetting a sparse subset of the parameters. The second one, Weight Shuffling, introduces an entropy- and weight distribution-invariant non-white noise to the parameters. The latter can also be interpreted as an ensemble approach. The proposed methods are evaluated on benchmark datasets, such as MNIST, CIFAR-10 or the JSB Chorales database, and also on time series modeling tasks. We report gains both regarding performance and entropy of the analyzed networks. We also made our code available as a GitHub repository (https://github.com/rpatrik96/lod-wmm-2019).

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

rpatrik96/lod-wmm-2019
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.