Block-Based Compression and Corresponding Hardware Circuits for Sparse Activations-Reference-Cited by-同舟云学术

Block-Based Compression and Corresponding Hardware Circuits for Sparse Activations

Published:2021-11-10 Issue:22 Volume:21 Page:7468
ISSN:1424-8220
Container-title:Sensors
language:en
Short-container-title:Sensors

Author:

Weng Yui-Kai^ORCID,Huang Shih-Hsu^ORCID,Kao Hsu-Yu^ORCID

Abstract

In a CNN (convolutional neural network) accelerator, to reduce memory traffic and power consumption, there is a need to exploit the sparsity of activation values. Therefore, some research efforts have been paid to skip ineffectual computations (i.e., multiplications by zero). Different from previous works, in this paper, we point out the similarity of activation values: (1) in the same layer of a CNN model, most feature maps are either highly dense or highly sparse; (2) in the same layer of a CNN model, feature maps in different channels are often similar. Based on the two observations, we propose a block-based compression approach, which utilizes both the sparsity and the similarity of activation values to further reduce the data volume. Moreover, we also design an encoder, a decoder and an indexing module to support the proposed approach. The encoder is used to translate output activations into the proposed block-based compression format, while both the decoder and the indexing module are used to align nonzero values for effectual computations. Compared with previous works, benchmark data consistently show that the proposed approach can greatly reduce both memory traffic and power consumption.

Funder

Ministry of Science and Technology, Taiwan

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Biochemistry,Instrumentation,Atomic and Molecular Physics, and Optics,Analytical Chemistry

Link

https://www.mdpi.com/1424-8220/21/22/7468/pdf

Reference32 articles.

1. A review on deep convolutional neural networks

2. Efficient Processing of Deep Neural Networks: A Tutorial and Survey

3. Deep learning for decentralized parking lot occupancy detection

4. Classification and Segmentation of Longitudinal Road Marking Using Convolutional Neural Networks for Dynamic Retroreflection Estimation

5. Deep CNN-Based Super-Resolution Using External and Internal Examples

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Dataflow and Hardware Design for The Sharing of Feature Maps;2022 19th International SoC Design Conference (ISOCC);2022-10-19

2. Network Pruning by Feature Map Sharing with K-Means Clustering;2022 IEEE International Conference on Consumer Electronics - Taiwan;2022-07-06