Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
Dream2Reward: Transition-Alignment Reward Models from Positive Demonstrations for Robotic Manipulation
Haoyu Zhang, Zecui Zeng, Bin Wang +3
Learning robotic policies requires dense rewards that remain informative when behavior departs from successful demonstrations. Progress-based rewards estimate how far an observatio…
cs.RO2025
Boosting Action-Information via a Variational Bottleneck on Unlabelled Robot Videos
Haoyu Zhang, Long Cheng
Learning from demonstrations (LfD) typically relies on large amounts of action-labeled expert trajectories, which fundamentally constrains the scale of available training data. A p…