상세 보기
Split Computing for Mobile Devices: Energy and Latency Perspective
- Jung, Daeyoung;
- Lee, Jaewook;
- Jeong, Hyeonjae;
- Cha, Dongju;
- Kim, Heewon;
- ... Pack, Sangheon
WEB OF SCIENCE
4SCOPUS
5초록
To tackle the difficulties of running sophisticated deep neural network (DNN) models on mobile devices, split computing presents a viable solution by offloading computations to the edge server. Current split computing schemes typically aim to lower either inference latency or energy use separately; however, optimizing both simultaneously is quite challenging due to numerous shifting factors, such as intensive continuous DNN model inferences, DNN model traits, and device/network conditions. Moreover, in practical applications, edge server overload might lead to substantial queuing delays, adding complexity to the optimization process. This article outlines a joint optimization problem that simultaneously seeks to minimize both inference latency and energy consumption, with a distinct inclusion of queue clearance latency for an accurate analysis of the continuously generated DNN model inferences. To address this intricate optimization challenge, we introduce a low-complexity heuristic algorithm that sets split point decisions based on the residual energy of mobile devices for each DNN inference cycle. Upon evaluation, our proposed algorithm demonstrates notable improvements by reducing inference latency by between 73.37% and 99.39%, and cutting down energy usage by between 39.97% and 94.67% compared to fully local processing on mobile devices.
키워드
- 제목
- Split Computing for Mobile Devices: Energy and Latency Perspective
- 저자
- Jung, Daeyoung; Lee, Jaewook; Jeong, Hyeonjae; Cha, Dongju; Kim, Heewon; Pack, Sangheon
- 발행일
- 2025-05
- 유형
- Article
- 권
- 18
- 호
- 3
- 페이지
- 1798 ~ 1810