7 citations · 7 across the 2 of their papers we have counts for
3 papers
Provable Defense against Backdoor Policies in Reinforcement Learning
Shubham Kumar Bharti, Xuezhou Zhang, Adish Singla +1
We propose a provable defense mechanism against backdoor policies in reinforcement learning under subspace trigger assumption. A backdoor policy is a security threat where an adver…
The Sample Complexity of Teaching-by-Reinforcement on Q-Learning
Xuezhou Zhang, Shubham Kumar Bharti, Yuzhe Ma +2
We study the sample complexity of teaching, termed as "teaching dimension" (TDim) in the literature, for the teaching-by-reinforcement paradigm, where the teacher guides the studen…
On the relationship between multitask neural networks and multitask Gaussian Processes
Karthikeyan K, Shubham Kumar Bharti, Piyush Rai
Despite the effectiveness of multitask deep neural network (MTDNN), there is a limited theoretical understanding on how the information is shared across different tasks in MTDNN. I…