102 citations · 254 across the 57 of their papers we have counts for
3 papers · 1 filter
Provable Defense against Backdoor Policies in Reinforcement Learning
Shubham Kumar Bharti, Xuezhou Zhang, Adish Singla +1
We propose a provable defense mechanism against backdoor policies in reinforcement learning under subspace trigger assumption. A backdoor policy is a security threat where an adver…
Equity and Fairness of Bayesian Knowledge Tracing
Sebastian Tschiatschek, Maria Knobelsdorf, Adish Singla
We consider the equity and fairness of curricula derived from Knowledge Tracing models. We begin by defining a unifying notion of an equitable tutoring system as a system that achi…
Admissible Policy Teaching through Reward Design
Kiarash Banihashem, Adish Singla, Jiarui Gan +1
We study reward design strategies for incentivizing a reinforcement learning agent to adopt a policy from a set of admissible policies. The goal of the reward designer is to modify…