4 papers · 1 filter
Risk-Aware Reinforcement Learning with Bandit-Based Adaptation for Quadrupedal Locomotion
Yuanhong Zeng, Anushri Dixit
In this work, we study risk-aware reinforcement learning for quadrupedal locomotion. Our approach trains a family of risk-conditioned policies using a Conditional Value-at-Risk (CV…
Reliable and Scalable Robot Policy Evaluation with Imperfect Simulators
Apurva Badithela, David Snyder, Lihan Zha +4
Rapid progress in imitation learning, foundation models, and large-scale datasets has led to robot manipulation policies that generalize to a wide-range of tasks and environments.…
Perceive With Confidence: Statistical Safety Assurances for Navigation with Learning-Based Perception
Zhiting Mei, Anushri Dixit, Meghan Booker +5
Rapid advances in perception have enabled large pre-trained models to be used out of the box for transforming high-dimensional, noisy, and partial observations of the world into ri…
Explore until Confident: Efficient Exploration for Embodied Question Answering
Allen Z. Ren, Jaden Clark, Anushri Dixit +3
We consider the problem of Embodied Question Answering (EQA), which refers to settings where an embodied agent such as a robot needs to actively explore an environment to gather in…