5 papers
Joint Optimization of Offloading, Batching and DVFS for Multiuser Co-Inference
Yaodan Xu, Sheng Zhou, Zhisheng Niu
With the growing integration of artificial intelligence in mobile applications, a substantial number of deep neural network (DNN) inference requests are generated daily by mobile d…
Optimization of Layer Skipping and Frequency Scaling for Convolutional Neural Networks under Latency Constraint
Minh David Thao Chan, Ruoyu Zhao, Yukuan Jia +2
The energy consumption of Convolutional Neural Networks (CNNs) is a critical factor in deploying deep learning models on resource-limited equipment such as mobile devices and auton…
SMDP-Based Dynamic Batching for Improving Responsiveness and Energy Efficiency of Batch Services
Yaodan Xu, Sheng Zhou, Zhisheng Niu
For servers incorporating parallel computing resources, batching is a pivotal technique for providing efficient and economical services at scale. Parallel computing resources exhib…
AEPHORA: AI/ML-Based Energy-Efficient Proactive Handover and Resource Allocation
Bowen Xie, Sheng Zhou, Zhisheng Niu +2
Future Vehicle-to-Everything (V2X) scenarios require high-speed, low-latency, and ultra-reliable communication services, particularly for applications such as autonomous driving an…
DiffCP: Ultra-Low Bit Collaborative Perception via Diffusion Model
Ruiqing Mao, Haotian Wu, Yukuan Jia +5
Collaborative perception (CP) is emerging as a promising solution to the inherent limitations of stand-alone intelligence. However, current wireless communication systems are unabl…