3 papers
cs.RO2026
InstructMove: A Text-Indispensable Benchmark for Instruction-Following Manipulation
Mengao Zhao, Ziang Li, Chaodong Huang +15
Vision-language-action (VLA) models have made general-purpose robot manipulation increasingly plausible by conditioning robot actions on natural-language instructions. A key test o…
cs.RO2026
EmbodiedGen V2: An Agentic, Simulation-Ready 3D World Engine for Embodied AI
Xinjie Wang, Liu Liu, Taojun Ding +9
We present EmbodiedGen V2, a generative 3D world engine for building executable policy-ready environments for embodied intelligence. Sim-ready 3D asset generation has advanced rapi…
cs.CV2025
ReCamDriving: LiDAR-Free Camera-Controlled Video Synthesis for Novel Trajectories
Yaokun Li, Shuaixian Wang, Mantang Guo +6
Synthesizing multi-pass videos is important for autonomous driving. While current repair-based methods often struggle with out-of-distribution artifacts, camera-controlled methods…