3 papers
cs.LG2025
Joint Memory Frequency and Computing Frequency Scaling for Energy-efficient DNN Inference
Yunchu Han, Zhaojun Nan, Sheng Zhou +1
Deep neural networks (DNNs) have been widely applied in diverse applications, but the problems of high latency and energy overhead are inevitable on resource-constrained devices. T…
cs.DC2025
Robust DNN Partitioning and Resource Allocation Under Uncertain Inference Time
Zhaojun Nan, Yunchu Han, Sheng Zhou +1
In edge intelligence systems, deep neural network (DNN) partitioning and data offloading can provide real-time task inference for resource-constrained mobile devices. However, the…
cs.LG2025
DVFS-Aware DNN Inference on GPUs: Latency Modeling and Performance Analysis
Yunchu Han, Zhaojun Nan, Sheng Zhou +1
The rapid development of deep neural networks (DNNs) is inherently accompanied by the problem of high computational costs. To tackle this challenge, dynamic voltage frequency scali…