2 papers
cs.AI2025
MiCRo: Mixture Modeling and Context-aware Routing for Personalized Preference Learning
Jingyan Shen, Jiarui Yao, Rui Yang +5
Reward modeling is a key step in building safe foundation models when applying reinforcement learning from human feedback (RLHF) to align Large Language Models (LLMs). However, rew…
cs.LG2024
2D-OOB: Attributing Data Contribution Through Joint Valuation Framework
Yifan Sun, Jingyan Shen, Yongchan Kwon
Data valuation has emerged as a powerful framework for quantifying each datum's contribution to the training of a machine learning model. However, it is crucial to recognize that t…