12 papers
Map-Det3D: Metric Feed-Forward 3D Reconstruction Prior for Multi-view 3D Object Detection from Streaming Inputs
Yung-Hsu Yang, Luigi Piccinelli, Samuel Rota Bulò +7
Metric 3D object detection is a core capability for embodied agents, yet most reliable systems lean on depth sensors, trading away cost, power, and integration simplicity. This mot…
LuxRemix: Lighting Decomposition and Remixing for Indoor Scenes
Ruofan Liang, Norman Müller, Ethan Weber +4
We present a novel approach for interactive light editing in indoor scenes from a single multi-view scene capture. Our method leverages a generative image-based light decomposition…
Learn2Splat: Extending the Horizon of Learned 3DGS Optimization
Naama Pearl, Stefano Esposito, Haofei Xu +6
3D Gaussian Splatting (3DGS) optimization is most commonly performed using standard optimizers (Adam, SGD). While stable across diverse scenes, standard optimizers are general-purp…
BulletGen: Improving 4D Reconstruction with Bullet-Time Generation
Denis Rozumny, Jonathon Luiten, Numair Khan +2
Transforming casually captured, monocular videos into fully immersive dynamic experiences is a highly ill-posed task, and comes with significant challenges, e.g., reconstructing un…
DRoPS: Dynamic 3D Reconstruction of Pre-Scanned Objects
Narek Tumanyan, Samuel Rota Bulò, Denis Rozumny +5
Dynamic scene reconstruction from casual videos has seen recent remarkable progress. Numerous approaches have attempted to overcome the ill-posedness of the task by distilling prio…
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Nikhil Keetha, Norman Müller, Johannes Schönberger +14
We introduce MapAnything, a unified transformer-based feed-forward model that ingests one or more images along with optional geometric inputs such as camera intrinsics, poses, dept…