SAC: An Ultra-Efficient Spin-based Architecture for Compressed DNNs-Reference-Cited by-同舟云学术

SAC: An Ultra-Efficient Spin-based Architecture for Compressed DNNs

Published:2024-01-19 Issue:1 Volume:21 Page:1-26
ISSN:1544-3566
Container-title:ACM Transactions on Architecture and Code Optimization
language:en
Short-container-title:ACM Trans. Archit. Code Optim.

Author:

Zhao Yunping¹^ORCID,Ma Sheng²^ORCID,Liu Heng³^ORCID,Huang Libo³^ORCID,Dai Yi³^ORCID

Affiliation:

1. Institute of Microelectronics and Microprocessors, School of Computy

2. The Science and Technology on Parallel and Distributed Processing Laboratory, National University of Defense Technology

3. School of Computer, National University of Defense Technology

Abstract

Deep Neural Networks (DNNs) have achieved great progress in academia and industry. But they have become computational and memory intensive with the increase of network depth. Previous designs seek breakthroughs in software and hardware levels to mitigate these challenges. At the software level, neural network compression techniques have effectively reduced network scale and energy consumption. However, the conventional compression algorithm is complex and energy intensive. At the hardware level, the improvements in the semiconductor process have effectively reduced power and energy consumption. However, it is difficult for the traditional Von-Neumann architecture to further reduce the power consumption, due to the memory wall and the end of Moore’s law. To overcome these challenges, the spintronic device based DNN machines have emerged for their non-volatility, ultra low power, and high energy efficiency. However, there is no spin-based design that has achieved innovation at both the software and hardware level. Specifically, there is no systematic study of spin-based DNN architecture to deploy compressed networks. In our study, we present an ultra-efficient Spin-based Architecture for Compressed DNNs (SAC), to substantially reduce power consumption and energy consumption. Specifically, we propose a One-Step Compression algorithm (OSC) to reduce the computational complexity with minimum accuracy loss. We also propose a spin-based architecture to realize better performance for the compressed network. Furthermore, we introduce a novel computation flow that enables the reuse of activations and weights. Experimental results show that our study can reduce the computational complexity of compression algorithm from 𝒪( Tk 3 to 𝒪( k 2 log k ), and achieve 14× ∼ 40× compression ratio. Furthermore, our design can attain a 2× enhancement in power efficiency and a 5× improvement in computational efficiency compared to the Eyeriss. Our models are available at an anonymous link https://bit.ly/39cdtTa .

Funder

National Key R&D

NSFC

NSF of Hunan Province

STIP of Hunan Province

Key Laboratory of Advanced Microprocessor Chips and Systems

Hunan Postgraduate Research Innovation

Publisher

Association for Computing Machinery (ACM)

Subject

Hardware and Architecture,Information Systems,Software

Link

https://dl.acm.org/doi/pdf/10.1145/3632957

Reference75 articles.

1. Martín Abadi. 2016. TensorFlow: Learning functions at scale. In Proceedings of the 21st ACM SIGPLAN International Conference on Functional Programming. 1–1.

2. Md. Hasibul Amin, Mohammed Elbtity, Mohammadreza Mohammadi, and Ramtin Zand. 2022. MRAM-based analog sigmoid function for in-memory computing. In Proceedings of the Great Lakes Symposium on VLSI 2022. 319–323.

3. Aayush Ankit, Abhronil Sengupta, Priyadarshini Panda, and Kaushik Roy. 2017. RESPARC: A reconfigurable and energy-efficient architecture with memristive crossbars for deep spiking neural networks. In 2017 54th ACM/EDAC/IEEE Design Automation Conference (DAC’17). 1–6.

4. Neuromorphic computing using non-volatile memory;Burr Geoffrey W.;Advances in Physics: X,2017

5. Eyeriss: An energy-efficient reconfigurable accelerator for deep convolutional neural networks;Chen Yu-Hsin;IEEE Journal of Solid-state Circuits,2016

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. SAL: Optimizing the Dataflow of Spin-based Architectures for Lightweight Neural Networks;ACM Transactions on Architecture and Code Optimization;2024-09-14

2. EPHA: An Energy-efficient Parallel Hybrid Architecture for ANNs and SNNs;ACM Transactions on Design Automation of Electronic Systems;2024-03-14