Reducing the model order of deep neural networks using information theory

Ming Tu; Visar Berisha; Yu Cao; Jae-sun Seo

doi:10.1109/ISVLSI.2016.117

Reducing the model order of deep neural networks using information theory

Ming Tu, Visar Berisha, Yu Cao, Jae-sun Seo

Research output: Chapter in Book/Report/Conference proceeding › Conference contribution

4 Scopus citations

Abstract

Deep neural networks are typically represented by a much larger number of parameters than shallow models, making them prohibitive for small footprint devices. Recent research shows that there is considerable redundancy in the parameter space of deep neural networks. In this paper, we propose a method to compress deep neural networks by using the Fisher Information metric, which we estimate through a stochastic optimization method that keeps track of second-order information in the network. We first remove unimportant parameters and then use non-uniform fixed point quantization to assign more bits to parameters with higher Fisher Information estimates. We evaluate our method on a classification task with a convolutional neural network trained on the MNIST data set. Experimental results show that our method outperforms existing methods for both network pruning and quantization.

Original language	English (US)
Title of host publication	Proceedings - IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016
Publisher	IEEE Computer Society
Pages	93-98
Number of pages	6
ISBN (Electronic)	9781467390385
DOIs	https://doi.org/10.1109/ISVLSI.2016.117
State	Published - Sep 2 2016
Event	15th IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016 - Pittsburgh, United States Duration: Jul 11 2016 → Jul 13 2016

Publication series

Name	Proceedings of IEEE Computer Society Annual Symposium on VLSI, ISVLSI
Volume	2016-September
ISSN (Print)	2159-3469
ISSN (Electronic)	2159-3477

Other

Other	15th IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016
Country/Territory	United States
City	Pittsburgh
Period	7/11/16 → 7/13/16

ASJC Scopus subject areas

Hardware and Architecture
Control and Systems Engineering
Electrical and Electronic Engineering

Access to Document

10.1109/ISVLSI.2016.117

Cite this

Tu, M., Berisha, V., Cao, Y., & Seo, J. (2016). Reducing the model order of deep neural networks using information theory. In Proceedings - IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016 (pp. 93-98). Article 7560179 (Proceedings of IEEE Computer Society Annual Symposium on VLSI, ISVLSI; Vol. 2016-September). IEEE Computer Society. https://doi.org/10.1109/ISVLSI.2016.117

Reducing the model order of deep neural networks using information theory. / Tu, Ming; Berisha, Visar; Cao, Yu et al.
Proceedings - IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016. IEEE Computer Society, 2016. p. 93-98 7560179 (Proceedings of IEEE Computer Society Annual Symposium on VLSI, ISVLSI; Vol. 2016-September).

Research output: Chapter in Book/Report/Conference proceeding › Conference contribution

Tu, M, Berisha, V, Cao, Y & Seo, J 2016, Reducing the model order of deep neural networks using information theory. in Proceedings - IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016., 7560179, Proceedings of IEEE Computer Society Annual Symposium on VLSI, ISVLSI, vol. 2016-September, IEEE Computer Society, pp. 93-98, 15th IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016, Pittsburgh, United States, 7/11/16. https://doi.org/10.1109/ISVLSI.2016.117

@inproceedings{02affbca454c4593989cfb269fe9a647,

title = "Reducing the model order of deep neural networks using information theory",

abstract = "Deep neural networks are typically represented by a much larger number of parameters than shallow models, making them prohibitive for small footprint devices. Recent research shows that there is considerable redundancy in the parameter space of deep neural networks. In this paper, we propose a method to compress deep neural networks by using the Fisher Information metric, which we estimate through a stochastic optimization method that keeps track of second-order information in the network. We first remove unimportant parameters and then use non-uniform fixed point quantization to assign more bits to parameters with higher Fisher Information estimates. We evaluate our method on a classification task with a convolutional neural network trained on the MNIST data set. Experimental results show that our method outperforms existing methods for both network pruning and quantization.",

author = "Ming Tu and Visar Berisha and Yu Cao and Jae-sun Seo",

note = "Publisher Copyright: {\textcopyright} 2016 IEEE.; 15th IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016 ; Conference date: 11-07-2016 Through 13-07-2016",

year = "2016",

month = sep,

day = "2",

doi = "10.1109/ISVLSI.2016.117",

language = "English (US)",

series = "Proceedings of IEEE Computer Society Annual Symposium on VLSI, ISVLSI",

publisher = "IEEE Computer Society",

pages = "93--98",

booktitle = "Proceedings - IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016",

}

TY - GEN

T1 - Reducing the model order of deep neural networks using information theory

AU - Tu, Ming

AU - Berisha, Visar

AU - Cao, Yu

AU - Seo, Jae-sun

PY - 2016/9/2

Y1 - 2016/9/2

N2 - Deep neural networks are typically represented by a much larger number of parameters than shallow models, making them prohibitive for small footprint devices. Recent research shows that there is considerable redundancy in the parameter space of deep neural networks. In this paper, we propose a method to compress deep neural networks by using the Fisher Information metric, which we estimate through a stochastic optimization method that keeps track of second-order information in the network. We first remove unimportant parameters and then use non-uniform fixed point quantization to assign more bits to parameters with higher Fisher Information estimates. We evaluate our method on a classification task with a convolutional neural network trained on the MNIST data set. Experimental results show that our method outperforms existing methods for both network pruning and quantization.

AB - Deep neural networks are typically represented by a much larger number of parameters than shallow models, making them prohibitive for small footprint devices. Recent research shows that there is considerable redundancy in the parameter space of deep neural networks. In this paper, we propose a method to compress deep neural networks by using the Fisher Information metric, which we estimate through a stochastic optimization method that keeps track of second-order information in the network. We first remove unimportant parameters and then use non-uniform fixed point quantization to assign more bits to parameters with higher Fisher Information estimates. We evaluate our method on a classification task with a convolutional neural network trained on the MNIST data set. Experimental results show that our method outperforms existing methods for both network pruning and quantization.

UR - http://www.scopus.com/inward/record.url?scp=84988929256&partnerID=8YFLogxK

UR - http://www.scopus.com/inward/citedby.url?scp=84988929256&partnerID=8YFLogxK

U2 - 10.1109/ISVLSI.2016.117

DO - 10.1109/ISVLSI.2016.117

M3 - Conference contribution

AN - SCOPUS:84988929256

T3 - Proceedings of IEEE Computer Society Annual Symposium on VLSI, ISVLSI

SP - 93

EP - 98

BT - Proceedings - IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016

PB - IEEE Computer Society

T2 - 15th IEEE Computer Society Annual Symposium on VLSI, ISVLSI 2016

Y2 - 11 July 2016 through 13 July 2016

ER -

Reducing the model order of deep neural networks using information theory

Abstract

Publication series

Other

ASJC Scopus subject areas

Access to Document

Other files and links

Fingerprint

Cite this