1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.RO2025
Learning Long-Context Diffusion Policies via Past-Token Prediction
Marcel Torne, Andy Tang, Yuejiang Liu +1
Reasoning over long sequences of observations and actions is essential for many robotic tasks. Yet, learning effective long-context policies from demonstrations remains challenging…
cs.LG2023★ 1 cited
Autonomous Robotic Reinforcement Learning with Asynchronous Human Feedback
Max Balsells, Marcel Torne, Zihan Wang +3
Ideally, we would place a robot in a real-world environment and leave it there improving on its own by gathering more experience autonomously. However, algorithms for autonomous ro…
cs.LG2023
Breadcrumbs to the Goal: Goal-Conditioned Exploration from Human-in-the-Loop Feedback
Marcel Torne, Max Balsells, Zihan Wang +4
Exploration and reward specification are fundamental and intertwined challenges for reinforcement learning. Solving sequential decision-making tasks requiring expansive exploration…