3 papers
cs.RO2026
J-LAW: Joint Localization and Actionable World Modeling via Coupled Latent Factor Graphs
Guanqun Cao, Liang Chen
Classical SLAM estimates metric poses and a geometric map but produces no actionable predictive model for planning. Action-conditioned world models learn compact latent dynamics fo…
cs.RO2025
Learn from the Past: Language-conditioned Object Rearrangement with Large Language Models
Guanqun Cao, Ryan Mckenna, Erich Graf +1
Object manipulation for rearrangement into a specific goal state is a significant task for collaborative robots. Accurately determining object placement is a key challenge, as misa…
cs.CV2025
Leveraging Stable Diffusion for Monocular Depth Estimation via Image Semantic Encoding
Jingming Xia, Guanqun Cao, Guang Ma +3
Monocular depth estimation involves predicting depth from a single RGB image and plays a crucial role in applications such as autonomous driving, robotic navigation, 3D reconstruct…