3 papers
cs.RO2025
Inference of Human-derived Specifications of Object Placement via Demonstration
Alex Cuellar, Ho Chit Siu, Julie A Shah
As robots' manipulation capabilities improve for pick-and-place tasks (e.g., object packing, sorting, and kitting), methods focused on understanding human-acceptable object configu…
cs.HC2025
In Pursuit of Predictive Models of Human Preferences Toward AI Teammates
Ho Chit Siu, Jaime D. Peña, Yutai Zhou +1
We seek measurable properties of AI agents that make them better or worse teammates from the subjective perspective of human collaborators. Our experiments use the cooperative card…
cs.LG2024
Accelerating Proximal Policy Optimization Learning Using Task Prediction for Solving Environments with Delayed Rewards
Ahmad Ahmad, Mehdi Kermanshah, Kevin Leahy +6
In this paper, we tackle the challenging problem of delayed rewards in reinforcement learning (RL). While Proximal Policy Optimization (PPO) has emerged as a leading Policy Gradien…