Reducing the Model Order of Deep Neural Networks Using Information   Theory

Ming Tu; Visar Berisha; Yu Cao; Jae-sun Seo

arXiv:1605.04859·cs.LG·May 17, 2016·1 cites

Reducing the Model Order of Deep Neural Networks Using Information Theory

Ming Tu, Visar Berisha, Yu Cao, Jae-sun Seo

PDF

Open Access

TL;DR

This paper introduces a novel neural network compression technique using Fisher Information to identify important parameters, enabling effective pruning and quantization, which improves model efficiency on small devices.

Contribution

The paper proposes a Fisher Information-based method for neural network compression that combines parameter pruning with non-uniform quantization, outperforming existing techniques.

Findings

01

Outperforms existing pruning and quantization methods

02

Effective reduction of model size with minimal accuracy loss

03

Validated on MNIST classification task

Abstract

Deep neural networks are typically represented by a much larger number of parameters than shallow models, making them prohibitive for small footprint devices. Recent research shows that there is considerable redundancy in the parameter space of deep neural networks. In this paper, we propose a method to compress deep neural networks by using the Fisher Information metric, which we estimate through a stochastic optimization method that keeps track of second-order information in the network. We first remove unimportant parameters and then use non-uniform fixed point quantization to assign more bits to parameters with higher Fisher Information estimates. We evaluate our method on a classification task with a convolutional neural network trained on the MNIST data set. Experimental results show that our method outperforms existing methods for both network pruning and quantization.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Neural Network Applications · Neural Networks and Applications · Anomaly Detection Techniques and Applications