Energy-Efficient Approximate Edge Inference Systems

被引:6
|
作者
Ghosh, Soumendu Kumar [1 ]
Raha, Arnab [2 ]
Raghunathan, Vijay [1 ]
机构
[1] Purdue Univ, Elmore Family Sch Elect & Comp Engn, 610 Purdue Mall, W Lafayette, IN 47907 USA
[2] Intel Corp, 2200 Mission Coll Blvd, Santa Clara, CA 95054 USA
关键词
Approximate computing; approximate systems; deep learning; DRAM; edge AI; edge-to-cloud computing; energy efficiency; quality-aware pruning; quality-energy tradeoff; CMOS IMAGE SENSOR; NEURAL-NETWORKS; PERFORMANCE; CHALLENGES;
D O I
10.1145/3589766
中图分类号
TP3 [计算技术、计算机技术];
学科分类号
0812 ;
摘要
The rapid proliferation of the Internet of Things and the dramatic resurgence of artificial intelligence based application workloads have led to immense interest in performing inference on energy-constrained edge devices. Approximate computing (a design paradigm that trades off a small degradation in application quality for disproportionate energy savings) is a promising technique to enable energy-efficient inference at the edge. This article introduces the concept of an approximate edge inference system (AxIS) and proposes a systematic methodology to perform joint approximations between different subsystems in a deep neural network (DNN)-based edge inference system, leading to significant energy benefits compared to approximating individual subsystems in isolation. We use a smart camera system that executes various DNN-based image classification and object detection applications to illustrate how the sensor, memory, compute, and communication subsystems can all be approximated synergistically. We demonstrate our proposed methodology using two variants of a smart camera system: (a) Cam(Edge), where the DNN is executed locally on the edge device, and (b) CamCloud, where the edge device sends the captured image to a remote cloud server that executes the DNN. We have prototyped such an approximate inference system using an Intel Stratix IV GX-based Terasic TR4-230 FPGA development board. Experimental results obtained using six large DNNs and four compact DNNs running image classification applications demonstrate significant energy savings (approximate to 1.6x-4.7x for large DNNs and approximate to 1.5x-3.6x for small DNNs), for minimal (<1%) loss in application-level quality. Furthermore, results using four object detection DNNs exhibit energy savings of approximate to 1.5x-5.2x for similar quality loss. Compared to approximating a single subsystem in isolation, AxIS achieves 1.05x-3.25x gains in energy savings for image classification and 1.35x-4.2x gains for object detection on average, for minimal (<1%) application-level quality loss.
引用
收藏
页数:50
相关论文
共 50 条
  • [41] Approximate Learning and Fault-Tolerant Mapping for Energy-Efficient Neuromorphic Systems
    Gebregirogis, Anteneh
    Tahoori, Mehdi
    ACM TRANSACTIONS ON DESIGN AUTOMATION OF ELECTRONIC SYSTEMS, 2021, 26 (03)
  • [42] Energy-Efficient Embedded Inference of SVMs on FPGA
    Elgawi, Osman
    Mutawa, A. M.
    Ahmad, Afaq
    2019 IEEE COMPUTER SOCIETY ANNUAL SYMPOSIUM ON VLSI (ISVLSI 2019), 2019, : 165 - 169
  • [43] Low-Power and Energy-Efficient Full Adders With Approximate Adiabatic Logic for Edge Computing
    Yang, Wu
    Thapliyal, Himanshu
    2020 IEEE COMPUTER SOCIETY ANNUAL SYMPOSIUM ON VLSI (ISVLSI 2020), 2020, : 312 - 315
  • [44] DVFO: Learning-Based DVFS for Energy-Efficient Edge-Cloud Collaborative Inference
    Zhang, Ziyang
    Zhao, Yang
    Li, Huan
    Lin, Changyao
    Liu, Jie
    IEEE TRANSACTIONS ON MOBILE COMPUTING, 2024, 23 (10) : 9042 - 9059
  • [45] Heterogeneous Memory Integration and Optimization for Energy-Efficient Multi-Task NLP Edge Inference
    Fu, Zirui
    Avaliani, Aleksandre
    Donato, Marco
    PROCEEDINGS OF THE 29TH ACM/IEEE INTERNATIONAL SYMPOSIUM ON LOW POWER ELECTRONICS AND DESIGN, ISLPED 2024, 2024,
  • [46] Energy-Efficient Joint Partitioning and Offloading for Delay-Sensitive CNN Inference in Edge Computing
    Zha, Zhiyong
    Yang, Yifei
    Xia, Yongjun
    Wang, Zhaoyi
    Luo, Bin
    Li, Kaihong
    Ye, Chenkai
    Xu, Bo
    Peng, Kai
    APPLIED SCIENCES-BASEL, 2024, 14 (19):
  • [47] EFFECT-DNN: Energy-efficient Edge Framework for Real-time DNN Inference
    Zhang, Xiaojie
    Mounesan, Motahare
    Debroy, Saptarshi
    2023 IEEE 24TH INTERNATIONAL SYMPOSIUM ON A WORLD OF WIRELESS, MOBILE AND MULTIMEDIA NETWORKS, WOWMOM, 2023, : 10 - 20
  • [48] Energy-efficient pumping systems
    Tutterow, VC
    Doolin, JH
    Paul, BO
    CHEMICAL PROCESSING, 1996, 59 (08): : 30 - &
  • [49] Energy-Efficient Task Caching and Offloading Strategy in Mobile Edge Computing Systems
    Chen, Qian
    Liu, Zhoubin
    Ruan, Linna
    Wang, Zixiang
    Shao, Sujie
    Qi, Feng
    SECURITY WITH INTELLIGENT COMPUTING AND BIG-DATA SERVICES, 2020, 895 : 824 - 837
  • [50] Energy-Efficient Task Offloading of Edge-Aided Maritime UAV Systems
    Li, Huanran
    Wu, Shaohua
    Jiao, Jian
    Lin, Xiao-Hui
    Zhang, Ning
    Zhang, Qinyu
    IEEE TRANSACTIONS ON VEHICULAR TECHNOLOGY, 2023, 72 (01) : 1116 - 1126