3 papers
cs.AI2026
S2T-RLHF: Hierarchical Credit Assignment for Stable Preference-Based RLHF
Wei Chen, Guanghui Zhu, Yafei Li +2
Reinforcement learning from human feedback (RLHF) with preference-based reward models often exhibits unstable training dynamics. A key contributing factor is that standard RLHF rel…
cs.LG2025
GEFM: Graph-Enhanced EEG Foundation Model
Limin Wang, Toyotaro Suzumura, Hiroki Kanezashi
Electroencephalography (EEG) signals provide critical insights for applications in disease diagnosis and healthcare. However, the scarcity of labeled EEG data poses a significant c…
cs.LG2024
SA-GNAS: Seed Architecture Expansion for Efficient Large-scale Graph Neural Architecture Search
Guanghui Zhu, Zipeng Ji, Jingyan Chen +3
GNAS (Graph Neural Architecture Search) has demonstrated great effectiveness in automatically designing the optimal graph neural architectures for multiple downstream tasks, such a…