9 papers
RayMap3R: Inference-Time RayMap for Dynamic 3D Reconstruction
Feiran Wang, Zezhou Shang, Gaowen Liu +1
Streaming feed-forward 3D reconstruction enables real-time joint estimation of scene geometry and camera poses from RGB images. However, without explicit dynamic reasoning, streami…
CogniMap3D: Cognitive 3D Mapping and Rapid Retrieval
Feiran Wang, Junyi Wu, Dawen Cai +2
We present CogniMap3D, a bioinspired framework for dynamic 3D scene understanding and reconstruction that emulates human cognitive processes. Our approach maintains a persistent me…
TraceFlow: Dynamic 3D Reconstruction of Specular Scenes Driven by Ray Tracing
Jiachen Tao, Junyi Wu, Haoxuan Wang +3
We present TraceFlow, a novel framework for high-fidelity rendering of dynamic specular scenes by addressing two key challenges: precise reflection direction estimation and physica…
GLaD: Geometric Latent Distillation for Vision-Language-Action Models
Minghao Guo, Meng Cao, Jiachen Tao +5
Most existing Vision-Language-Action (VLA) models rely primarily on RGB information, while ignoring geometric cues crucial for spatial reasoning and manipulation. In this work, we…
Motion Marionette: Rethinking Rigid Motion Transfer via Prior Guidance
Haoxuan Wang, Jiachen Tao, Junyi Wu +3
We present Motion Marionette, a zero-shot framework for rigid motion transfer from monocular source videos to single-view target images. Previous works typically employ geometric,…
Orientation-anchored Hyper-Gaussian for 4D Reconstruction from Casual Videos
Junyi Wu, Jiachen Tao, Haoxuan Wang +3
We present Orientation-anchored Gaussian Splatting (OriGS), a novel framework for high-quality 4D reconstruction from casually captured monocular videos. While recent advances exte…