DVFO: Learning-Based DVFS for Energy-Efficient Edge-Cloud Collaborative   Inference

Ziyang Zhang; Yang Zhao; Huan Li; Changyao Lin; and Jie Liu

arXiv:2306.01811·cs.LG·June 26, 2023·2 cites

DVFO: Learning-Based DVFS for Energy-Efficient Edge-Cloud Collaborative Inference

Ziyang Zhang, Yang Zhao, Huan Li, Changyao Lin, and Jie Liu

PDF

Open Access

TL;DR

DVFO is a deep reinforcement learning-based framework that co-optimizes DVFS and offloading to enhance energy efficiency and reduce latency in edge-cloud DNN inference.

Contribution

It introduces a novel DRL-based co-optimization of DVFS and offloading parameters for edge devices in DNN inference.

Findings

01

Reduces energy consumption by 33% on average.

02

Achieves up to 59.1% latency reduction.

03

Maintains accuracy within 1% loss.

Abstract

Due to limited resources on edge and different characteristics of deep neural network (DNN) models, it is a big challenge to optimize DNN inference performance in terms of energy consumption and end-to-end latency on edge devices. In addition to the dynamic voltage frequency scaling (DVFS) technique, the edge-cloud architecture provides a collaborative approach for efficient DNN inference. However, current edge-cloud collaborative inference methods have not optimized various compute resources on edge devices. Thus, we propose DVFO, a novel DVFS-enabled edge-cloud collaborative inference framework, which co-optimizes DVFS and offloading parameters via deep reinforcement learning (DRL). Specifically, DVFO automatically co-optimizes 1) the CPU, GPU and memory frequencies of edge devices, and 2) the feature maps to be offloaded to cloud servers. In addition, it leverages a…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsIoT and Edge/Fog Computing · Brain Tumor Detection and Classification · Advanced Neural Network Applications