Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Towards Reward Fairness in RLHF: From a Resource Allocation Perspective
Sheng Ouyang, Yulan Hu, Ge Chen +3
Rewards serve as proxies for human preferences and play a crucial role in Reinforcement Learning from Human Feedback (RLHF). However, if these rewards are inherently imperfect, exh…
cs.LG2024
Preserving Node Distinctness in Graph Autoencoders via Similarity Distillation
Ge Chen, Yulan Hu, Sheng Ouyang +2
Graph autoencoders (GAEs), as a kind of generative self-supervised learning approach, have shown great potential in recent years. GAEs typically rely on distance-based criteria, su…