3 papers
cs.RO2026
GIF: Agentic Generation of Interactive and Functional Object Compositions for Robot Learning
Long Xu, Zhiqi Zhang, Mi Yan +8
Robot manipulation foundation models require scalable evaluation and data generation across diverse scenarios, with simulation providing an environment for both. Automated scene ge…
cs.CV2026
ReplicateAnyScene: Zero-Shot Video-to-3D Composition via Textual-Visual-Spatial Alignment
Mingyu Dong, Chong Xia, Mingyuan Jia +4
Humans exhibit an innate capacity to rapidly perceive and segment objects from video observations, and even mentally assemble them into structured 3D scenes. Replicating such capab…
cs.CV2026
Revisiting 3D Reconstruction Kernels as Low-Pass Filters
Shengjun Zhang, Min Chen, Yibo Wei +2
3D reconstruction is to recover 3D signals from the sampled discrete 2D pixels, with the goal to converge continuous 3D spaces. In this paper, we revisit 3D reconstruction from the…