4 papers
Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models
Yating Wang, Wenting Zhao, Yaqi Zhao +3
Large language models store not only isolated facts but also rules that support reasoning across symbolic expressions, natural language explanations, and concrete instances. Yet mo…
Value-Guidance MeanFlow for Offline Multi-Agent Reinforcement Learning
Teng Pang, Zhiqiang Dong, Yan Zhang +3
Offline multi-agent reinforcement learning (MARL) aims to learn the optimal joint policy from pre-collected datasets, requiring a trade-off between maximizing global returns and mi…
Diffusion Classifier-Driven Reward for Offline Preference-based Reinforcement Learning
Teng Pang, Bingzheng Wang, Guoqiang Wu +1
Offline preference-based reinforcement learning (PbRL) mitigates the need for reward definition, aligning with human preferences via preference-driven reward feedback without inter…
Towards Macro-AUC oriented Imbalanced Multi-Label Continual Learning
Yan Zhang, Guoqiang Wu, Bingzheng Wang +3
In Continual Learning (CL), while existing work primarily focuses on the multi-class classification task, there has been limited research on Multi-Label Learning (MLL). In practice…