5 papers
ReFPO: Reflow Regularization for Flow Matching Policy Gradients
Ge Wang, Yibo Peng, Fan Feng +10
We present Reflow-regularized Flow Matching Policy Gradients (ReFPO), a simple online RL method that adds explicit Reflow regularization to FPO for efficient flow-based control. We…
RoboSeek: You Need to Interact with Your Objects
Yibo Peng, Jiahao Yang, Shenhao Yan +5
Optimizing and refining action execution through exploration and interaction is a promising way for robotic manipulation. However, practical approaches to interaction-driven roboti…
Empowering Large Language Models with 3D Situation Awareness
Zhihao Yuan, Yibo Peng, Jinke Ren +7
Driven by the great success of Large Language Models (LLMs) in the 2D image domain, their applications in 3D scene understanding has emerged as a new trend. A key difference betwee…
CLEA: Closed-Loop Embodied Agent for Enhancing Task Execution in Dynamic Environments
Mingcong Lei, Ge Wang, Yiming Zhao +7
Large Language Models (LLMs) exhibit remarkable capabilities in the hierarchical decomposition of complex tasks through semantic reasoning. However, their application in embodied s…
STMA: A Spatio-Temporal Memory Agent for Long-Horizon Embodied Task Planning
Mingcong Lei, Yiming Zhao, Ge Wang +4
A key objective of embodied intelligence is enabling agents to perform long-horizon tasks in dynamic environments while maintaining robust decision-making and adaptability. To achi…