3 papers
cs.CR2026
Trigger the Straggler: Load Hijack on Mixture-of-Experts LLMs
Rui Zhang, Wenbo Jiang, Hongwei Li +4
Expert parallelism (EP) is a common strategy for serving large Mixture-of-Experts (MoE) models across multiple GPUs by distributing experts among devices. Router decisions then det…
cs.LG2025
Secure Resource Allocation via Constrained Deep Reinforcement Learning
Jianfei Sun, Qiang Gao, Cong Wu +3
The proliferation of Internet of Things (IoT) devices and the advent of 6G technologies have introduced computationally intensive tasks that often surpass the processing capabiliti…
cs.LG2024
An Efficient Privacy-aware Split Learning Framework for Satellite Communications
Jianfei Sun, Cong Wu, Shahid Mumtaz +4
In the rapidly evolving domain of satellite communications, integrating advanced machine learning techniques, particularly split learning, is crucial for enhancing data processing…