Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Joint Memory Frequency and Computing Frequency Scaling for Energy-efficient DNN Inference
Yunchu Han, Zhaojun Nan, Sheng Zhou +1
Deep neural networks (DNNs) have been widely applied in diverse applications, but the problems of high latency and energy overhead are inevitable on resource-constrained devices. T…
cs.LG2025
DVFS-Aware DNN Inference on GPUs: Latency Modeling and Performance Analysis
Yunchu Han, Zhaojun Nan, Sheng Zhou +1
The rapid development of deep neural networks (DNNs) is inherently accompanied by the problem of high computational costs. To tackle this challenge, dynamic voltage frequency scali…