Structured pruning of recurrent neural networks through neuron selection.

Journal: Neural networks : the official journal of the International Neural Network Society

Published Date: Dec 5, 2019

Abstract

Recurrent neural networks (RNNs) have recently achieved remarkable successes in a number of applications. However, the huge sizes and computational burden of these models make it difficult for their deployment on edge devices. A practically effective approach is to reduce the overall storage and computation costs of RNNs by network pruning techniques. Despite their successful applications, those pruning methods based on Lasso either produce irregular sparse patterns in weight matrices, which is not helpful in practical speedup. To address these issues, we propose a structured pruning method through neuron selection which can remove the independent neuron of RNNs. More specifically, we introduce two sets of binary random variables, which can be interpreted as gates or switches to the input neurons and the hidden neurons, respectively. We demonstrate that the corresponding optimization problem can be addressed by minimizing the L norm of the weight matrix. Finally, experimental results on language modeling and machine reading comprehension tasks have indicated the advantages of the proposed method in comparison with state-of-the-art pruning competitors. In particular, nearly 20× practical speedup during inference was achieved without losing performance for the language model on the Penn TreeBank dataset, indicating the promising performance of the proposed method.

Authors

Liangjian Wen

SMILE Lab, School of Computer Science and Engineering, University of Electronic Science and Technology of China, Chengdu 610031, China.
Xuanyang Zhang

SMILE Lab, School of Computer Science and Engineering, University of Electronic Science and Technology of China, Chengdu 610031, China.
Haoli Bai

Department of Computer Science and Engineering, The Chinese University of Hong Kong, Shatin NT 999077, Hong Kong SAR.
Zenglin Xu

Big Data Research Center, University of Electronic Science & Technology, Chengdu, Sichuan, China; School of Computer Science and Engineering, University of Electronic Science & Technology, Chengdu, Sichuan, China. Electronic address: zlxu@uestc.edu.cn.

Keywords

Data Compression Natural Language Processing Neural Networks, Computer

External Resources

View on PubMed Access via DOI PubMed (31855748)

Structured pruning of recurrent neural networks through neuron selection.

Abstract

Authors

Keywords

External Resources

Popular Topics

Recent Journals