8 papers
MVP-Nav: Multi-layer Value Map Planner Navigator
Wenyuan Xie, Shaokai Wu, Yijin Zhou +7
Zero-shot Object Goal Navigation (ZSON) with RGB-only perception poses a fundamental challenge for embodied agents, as the absence of explicit depth information introduces severe p…
UniTac: A Unified Multimodal Model for Cross-Sensor Tactile Understanding and Generation
Jiahang Tu, Fengyu Yang, Chenyang Ma +8
Unified multimodal models (UMMs) have shown great promise in integrating understanding and generation across diverse modalities. However, existing research rarely extends this para…
RelAfford6D: Relational 6D Affordance Graphs for Constraint-Driven Robotic Manipulation
Guodong Zhang, Qichen He, Wenyuan Xie +6
Bridging abstract semantics and precise physical control remains a fundamental challenge in open-world robotic manipulation. While recent data-driven policies show promise, their r…
Recovering Hidden Reward in Diffusion-Based Policies
Yanbiao Ji, Qiuchang Li, Yuting Hu +7
This paper introduces EnergyFlow, a framework that unifies generative action modeling with inverse reinforcement learning by parameterizing a scalar energy function whose gradient…
Dejavu: Towards Experience Feedback Learning for Embodied Intelligence
Shaokai Wu, Yanbiao Ji, Qiuchang Li +7
Embodied agents face a fundamental limitation: once deployed in real-world environments, they cannot easily acquire new knowledge to improve task performance. In this paper, we pro…
MORE: Multi-Organ Medical Image REconstruction Dataset
Shaokai Wu, Yapan Guo, Yanbiao Ji +6
CT reconstruction provides radiologists with images for diagnosis and treatment, yet current deep learning methods are typically limited to specific anatomies and datasets, hinderi…