From the 1 of 4 linked papers with an AI index.
4 papers
Training and Evaluating Ethical Reinforcement Learning Agents on Per-Episode Distributions
Prabhjyot Singh, Majid Ghasemi, Mark Crowley
Reinforcement Learning (RL) agents trained on a single reward signal exploit the gap between the designed reward and the intended behavior. This is particularly a problem when we a…
Learning When to Trust in Contextual Social Bandits
Majid Ghasemi, Mark Crowley
The paper studies bandit learning where feedback providers are honest in some contexts but biased in others (contextual sycophancy) and shows that without occasional ground‑truth a…
Preference-based Antibody Expression Ranking: Scaling with Large-scale Weak Supervision
Josh Qixuan Sun, Morteza Babaie, Wenyang Hou +2
Antibody expression ranking is a critical task in antibody design, yet its modelling is severely hindered by the scarcity of labeled experimental data. To address this, we propose…
Objective Decoupling in Social Reinforcement Learning: Recovering Ground Truth from Sycophantic Majorities
Majid Ghasemi, Mark Crowley
Contemporary AI alignment strategies rely on a fragile premise: that human feedback, while noisy, remains a fundamentally truthful signal. In this paper, we identify this assumptio…