Energy-Efficient Approximate Edge Inference Systems
The rapid proliferation of the Internet of Things (IoT) and the dramatic resurgence of artificial intelligence (AI) based application workloads have led to immense interest in performing inference on energy-constrained edge devices. Approximate computing (a design paradigm that trades off a small degradation in application quality for disproportionate energy savings) is a promising technique to enable energy-efficient inference at the edge. This paper introduces the concept of an approximate edge inference system (
AxIS
) and proposes a systematic methodology to perform joint approximations between different subsystems in a deep neural network (DNN)-based edge inference system, leading to significant energy benefits compared to approximating individual subsystems in isolation. We use a smart camera system that executes various DNN-based image classification and object detection applications to illustrate how the sensor, memory, compute, and communication subsystems can all be approximated synergistically. We demonstrate our proposed methodology using two variants of a smart camera system: (a)
Cam
Edge
, where the DNN is executed locally on the edge device, and (b)
Cam
Cloud
, where the edge device sends the captured image to a remote cloud server that executes the DNN. We have prototyped such an approximate inference system using an Intel Stratix IV GX-based Terasic TR4-230 FPGA development board. Experimental results obtained using six large DNNs and four compact DNNs running image classification applications demonstrate significant energy savings (≈ 1.6 ×–4.7 × for large DNNs and ≈ 1.5 ×–3.6 × for small DNNs) for minimal (<
Top-30
Journals
|
1
|
|
|
Transactions on Embedded Computing Systems
1 publication, 7.14%
|
|
|
IEEE Internet of Things Journal
1 publication, 7.14%
|
|
|
IEEE Access
1 publication, 7.14%
|
|
|
Lecture Notes in Computer Science
1 publication, 7.14%
|
|
|
ACM Computing Surveys
1 publication, 7.14%
|
|
|
IEEE Embedded Systems Letters
1 publication, 7.14%
|
|
|
Frontiers in High Performance Computing
1 publication, 7.14%
|
|
|
IEEE Transactions on Cloud Computing
1 publication, 7.14%
|
|
|
ACM Transactions on Design Automation of Electronic Systems
1 publication, 7.14%
|
|
|
1
|
Publishers
|
1
2
3
4
5
6
7
8
|
|
|
Institute of Electrical and Electronics Engineers (IEEE)
8 publications, 57.14%
|
|
|
Association for Computing Machinery (ACM)
4 publications, 28.57%
|
|
|
Springer Nature
1 publication, 7.14%
|
|
|
Frontiers Media S.A.
1 publication, 7.14%
|
|
|
1
2
3
4
5
6
7
8
|
- We do not take into account publications without a DOI.
- Statistics recalculated weekly.