2 papers
cs.RO2026
Risk-Aware Preference Learning for Stochastic Outcomes
Yi-Shiuan Tung, Yuni Wu, Wei Jiang +2
Learning reward functions from human preferences is a widely used approach for aligning robot behavior with user expectations in human-robot interaction. Most existing approaches a…
cs.RO2026
CRED: Counterfactual Reasoning and Environment Design for Active Preference Learning
Yi-Shiuan Tung, Gyanig Kumar, Wei Jiang +2
As a robot's operational environment and tasks to perform within it grow in complexity, the explicit specification and balancing of optimization objectives to achieve a preferred b…