3 papers
cs.LG2025
Capacity-Constrained Continual Learning
Zheng Wen, Doina Precup, Benjamin Van Roy +1
Any agents we can possibly build are subject to capacity constraints, as memory and compute resources are inherently finite. However, comparatively little attention has been dedica…
stat.ML2024
Reinforcement Learning in Credit Scoring and Underwriting
Seksan Kiatsupaibul, Pakawan Chansiripas, Pojtanut Manopanjasiri +2
This paper proposes a novel reinforcement learning (RL) framework for credit underwriting that tackles ungeneralizable contextual challenges. We adapt RL principles for credit scor…
cs.LG2024
RLHF and IIA: Perverse Incentives
Wanqiao Xu, Shi Dong, Xiuyuan Lu +3
Existing algorithms for reinforcement learning from human feedback (RLHF) can incentivize responses at odds with preferences because they are based on models that assume independen…