adversarial bias 1audit-based reinforcement learning 1contextual bandits 1regret analysis 1social feedback 1trust learning 1
From the 1 of 4 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Training and Evaluating Ethical Reinforcement Learning Agents on Per-Episode Distributions
Prabhjyot Singh, Majid Ghasemi, Mark Crowley
Reinforcement Learning (RL) agents trained on a single reward signal exploit the gap between the designed reward and the intended behavior. This is particularly a problem when we a…
cs.LG2026
Preference-based Antibody Expression Ranking: Scaling with Large-scale Weak Supervision
Josh Qixuan Sun, Morteza Babaie, Wenyang Hou +2
Antibody expression ranking is a critical task in antibody design, yet its modelling is severely hindered by the scarcity of labeled experimental data. To address this, we propose…