3 papers
cs.RO2026
Adaptive Querying for Reward Learning from Human Feedback
Yashwanthi Anand, Nnamdi Nwagwu, Kevin Sabbe +2
Learning from human feedback is a popular approach to train robots to adapt to user preferences and improve safety. Existing approaches typically consider a single querying (intera…
cs.AI2026
Uncovering Systemic and Environment Errors in Autonomous Systems Using Differential Testing
Yashwanthi Anand, Rahil P Mehta, Manish Motwani +1
When an autonomous agent behaves undesirably, including failure to complete a task, it can be difficult to determine whether the behavior is due to a systemic agent error, such as…
cs.AI2025
Multi-Objective Planning with Contextual Lexicographic Reward Preferences
Pulkit Rustagi, Yashwanthi Anand, Sandhya Saisubramanian
Autonomous agents are often required to plan under multiple objectives whose preference ordering varies based on context. The agent may encounter multiple contexts during its cours…