diffusion models 1human-in-the-loop 1policy adaptation 1reinforcement learning 1vision-language-action 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.RO2026
UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation
Junjie Lu, Xinyao Qin, Yuhua Jiang +6
The paper introduces UniSteer, a framework that converts human corrective actions into noise targets to guide a lightweight noise-prediction actor while simultaneously training it…
cs.RO2026
Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation
Xinyao Qin, Junjie Lu, Kaixin Wang +7
Human demonstrations for robot imitation learning often contain mistakes and corrective behaviors, such as imprecise grasps, object misalignment, unstable contact, and repeated att…
cs.AI2026
SEAGym: An Evaluation Environment for Self-Evolving LLM Agents
Congjie Zheng, Chuanyi Xue, Bin Liang +2
Self-evolving LLM-based agents improve mainly by changing their agent harness: the structured execution layer around a base model, including prompts, memory, tools, middleware, run…