3 papers
cs.CV2026
Real2Sim in HOI: Toward Physically Plausible HOI Reconstruction from Monocular Videos
Yubo Zhao, Yujin Chai, Yunao Dong +4
Recovering 4D human-object interaction (HOI) from monocular video is a key step toward scalable 3D content creation, embodied AI, and simulation-based learning. Recent methods can…
cs.CV2026
CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos
Chengfeng Zhao, Jiazhi Shu, Yubo Zhao +7
In this paper, we find that the generation of 3D human motions and 2D human videos is intrinsically coupled. 3D motions provide the structural prior for plausibility and consistenc…
cs.AI2025
Navigating Motion Agents in Dynamic and Cluttered Environments through LLM Reasoning
Yubo Zhao, Qi Wu, Yifan Wang +2
This paper advances motion agents empowered by large language models (LLMs) toward autonomous navigation in dynamic and cluttered environments, significantly surpassing first and r…