Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
FLEX-MoE: Federated Mixture-of-Experts with Load-balanced Expert Assignment for Edge Computing
Boyang Zhang, Xiaobing Chen, Songyang Zhang +4
Mixture-of-Experts (MoE) models enable scalable neural networks through conditional computation, offering enhanced effectiveness and efficiency for next-generation wireless communi…
cs.LG2025
Pruning and Malicious Injection: A Retraining-Free Backdoor Attack on Transformer Models
Taibiao Zhao, Mingxuan Sun, Hao Wang +2
Transformer models have demonstrated exceptional performance and have become indispensable in computer vision (CV) and natural language processing (NLP) tasks. However, recent stud…
cs.LG2025
Efficient Training of Large-Scale AI Models Through Federated Mixture-of-Experts: A System-Level Approach
Xiaobing Chen, Boyang Zhang, Xiangwei Zhou +4
The integration of Federated Learning (FL) and Mixture-of-Experts (MoE) presents a compelling pathway for training more powerful, large-scale artificial intelligence models (LAMs)…