4 papers
Token Communication for Multimodal Large Language Model
Jingkai Ying, Zhijin Qin, Yuan Shen +1
With the broad success of the Transformer architecture, token is becoming a new basic information processing unit. This trend is especially evident in multimodal large language mod…
MDGAM-Based Cooperative Task Scheduling for Communication-Constrained Distributed Multi-Agent Systems
Licheng Wang, Mingtao Huang, Yuan Shen
Cooperative task scheduling in communication-constrained distributed multi-agent systems is challenging because each agent must make decisions from partial and dynamic observations…
EGAM: Extended Graph Attention Model for Solving Routing Problems
Licheng Wang, Yuzi Yan, Mingtao Huang +1
Neural combinatorial optimization (NCO) solvers, implemented with graph neural networks (GNNs), have introduced new approaches for solving routing problems. Trained with reinforcem…
Reward-Robust RLHF in LLMs
Yuzi Yan, Xingzhou Lou, Jialian Li +6
As Large Language Models (LLMs) continue to progress toward more advanced forms of intelligence, Reinforcement Learning from Human Feedback (RLHF) is increasingly seen as a key pat…