Energy-Efficient Approximate Edge Inference Systems-Reference-Cited by-同舟云学术

Energy-Efficient Approximate Edge Inference Systems

Published:2023-07-24 Issue:4 Volume:22 Page:1-50
ISSN:1539-9087
Container-title:ACM Transactions on Embedded Computing Systems
language:en
Short-container-title:ACM Trans. Embed. Comput. Syst.

Author:

Ghosh Soumendu Kumar¹^ORCID,Raha Arnab²^ORCID,Raghunathan Vijay¹^ORCID

Affiliation:

1. Purdue University

2. Intel Corporation

Abstract

The rapid proliferation of the Internet of Things and the dramatic resurgence of artificial intelligence based application workloads have led to immense interest in performing inference on energy-constrained edge devices. Approximate computing (a design paradigm that trades off a small degradation in application quality for disproportionate energy savings) is a promising technique to enable energy-efficient inference at the edge. This article introduces the concept of an approximate edge inference system ( AxIS ) and proposes a systematic methodology to perform joint approximations between different subsystems in a deep neural network (DNN)-based edge inference system, leading to significant energy benefits compared to approximating individual subsystems in isolation. We use a smart camera system that executes various DNN-based image classification and object detection applications to illustrate how the sensor, memory, compute, and communication subsystems can all be approximated synergistically. We demonstrate our proposed methodology using two variants of a smart camera system: (a) Cam Edge , where the DNN is executed locally on the edge device, and (b) Cam Cloud , where the edge device sends the captured image to a remote cloud server that executes the DNN. We have prototyped such an approximate inference system using an Intel Stratix IV GX-based Terasic TR4-230 FPGA development board. Experimental results obtained using six large DNNs and four compact DNNs running image classification applications demonstrate significant energy savings (≈ 1.6× -4.7× for large DNNs and ≈ 1.5× -3.6× for small DNNs), for minimal (<1%) loss in application-level quality. Furthermore, results using four object detection DNNs exhibit energy savings of ≈ 1.5× -5.2× for similar quality loss. Compared to approximating a single subsystem in isolation, AxIS achieves 1.05× -3.25× gains in energy savings for image classification and 1.35× -4.2× gains for object detection on average, for minimal (<1%) application-level quality loss.

Funder

Center for Brain-inspired Computing Enabling Autonomous Intelligence

Joint University Microelectronics Program

Semiconductor Research Corporation

Defense Advanced Research Projects Agency

Publisher

Association for Computing Machinery (ACM)

Subject

Hardware and Architecture,Software

Link

https://dl.acm.org/doi/pdf/10.1145/3589766

Reference112 articles.

1. Less is more

2. Wild patterns: Ten years after the rise of adversarial machine learning

3. Exploiting approximate computing for deep learning acceleration

4. AdderNet: Do We Really Need Multiplications in Deep Learning?

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Toward Energy-Efficient Collaborative Inference Using Multisystem Approximations;IEEE Internet of Things Journal;2024-05-15

2. PArtNNer: Platform-Agnostic Adaptive Edge-Cloud DNN Partitioning for Minimizing End-to-End Latency;ACM Transactions on Embedded Computing Systems;2024-01-10

3. HARVEST: Towards Efficient Sparse DNN Accelerators using Programmable Thresholds;2024 37th International Conference on VLSI Design and 2024 23rd International Conference on Embedded Systems (VLSID);2024-01-06