6 citations · 12 across the 8 of their papers we have counts for
8 papers
FourierCompress: Layer-Aware Spectral Activation Compression for Efficient and Accurate Collaborative LLM Inference
Jian Ma, Xinchen Lyu, Jun Jiang +4
Collaborative large language model (LLM) inference enables real-time, privacy-preserving AI services on resource-constrained edge devices by partitioning computational workloads be…
Agile Orchestration at Will: An Entire Smart Service-Based Security Architecture Towards 6G
Zhuoran Duan, Guoshun Nan, Rushan Li +8
The upcoming 6G will fundamentally reshape mobile networks beyond communications, unlocking a multitude of applications that were once considered unimaginable. Meanwhile, security…
Convergence-Privacy-Fairness Trade-Off in Personalized Federated Learning
Xiyu Zhao, Qimei Cui, Weicai Li +5
Personalized federated learning (PFL), e.g., the renowned Ditto, strikes a balance between personalization and generalization by conducting federated learning (FL) to guide persona…
A Novel Indicator for Quantifying and Minimizing Information Utility Loss of Robot Teams
Xiyu Zhao, Qimei Cui, Wei Ni +5
The timely exchange of information among robots within a team is vital, but it can be constrained by limited wireless capacity. The inability to deliver information promptly can re…
Enhancing Convergence, Privacy and Fairness for Wireless Personalized Federated Learning: Quantization-Assisted Min-Max Fair Scheduling
Xiyu Zhao, Qimei Cui, Ziqiang Du +6
Personalized federated learning (PFL) offers a solution to balancing personalization and generalization by conducting federated learning (FL) to guide personalized learning (PL). L…
Two Is Better Than One: Rotations Scale LoRAs
Hongcan Guo, Guoshun Nan, Yuan Yang +9
Scaling Low-Rank Adaptation (LoRA)-based Mixture-of-Experts (MoE) facilitates large language models (LLMs) to efficiently adapt to diverse tasks. However, traditional gating mechan…