collaborators

11 papers

cs.RO2026

Difference-Aware Retrieval Policies for Imitation Learning

Quinn Pfeifer, Ethan Pronovost, Paarth Shah +3

Parametric imitation learning via behavior cloning can suffer from poor generalization to out-of-distribution states due to compounding errors during deployment. We show that reusi…

cs.LG2026

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning

Daniel Lawson, Adriana Hugessen, Charlotte Cloutier +2

While goal-conditioned behavior cloning (GCBC) methods can perform well on in-distribution training tasks, they do not necessarily generalize zero-shot to tasks that require condit…

cs.LG2026

Preventing Learning Stagnation in PPO by Scaling to 1 Million Parallel Environments

Michael Beukman, Khimya Khetarpal, Zeyu Zheng +4

An agent's performance stagnating at a suboptimal level is a common problem in deep on-policy RL. Focusing on PPO, we show that plateaus in certain regimes arise not because of kno…

cs.LG2026

Affordances Enable Partial World Modeling with LLMs

Khimya Khetarpal, Gheorghe Comanici, Jonathan Richens +5

Full models of the world require complex knowledge of immense detail. While pre-trained large models have been hypothesized to contain similar knowledge due to extensive pre-traini…

cs.LG2026

The Geometry of Learning to Avoid Interventions

Ethan Pronovost, Khimya Khetarpal, Siddhartha Srinivasa

Human interventions are a common source of supervision in autonomous systems during deployment. Many existing approaches are based on avoiding interventions, yet the consequences o…

cs.AI2025

Plasticity as the Mirror of Empowerment

David Abel, Michael Bowling, André Barreto +13

Agents are minimally entities that are influenced by their past observations and act to influence future observations. This latter capacity is captured by empowerment, which has se…