3d reconstruction 1feed-forward networks 1gaussian splatting 1geometry encoding 1monocular video 1multi-agent 1neural radiance fields 1pose graph optimization 1rgb-d 1semantic features 1slam 1
From the 2 of 4 linked papers with an AI index.
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
MAGiSt3R: Multi-Agent Feed-forward 3D Reconstruction from Monocular RGB Videos
Ziren Gong, Xiaohan Li, Fabio Tosi +4
The paper introduces MAGiSt3R, a multi-agent framework that reconstructs 3D scenes and tracks camera pose from monocular RGB videos in near real-time using feed-forward models and…
cs.CV2026
DINO-SLAM: DINO-informed RGB-D SLAM for Neural Implicit and Explicit Representations
Ziren Gong, Xiaohan Li, Fabio Tosi +4
The paper introduces DINO-SLAM, a system that combines DINO semantic features with a geometry encoder to improve both neural implicit (NeRF) and explicit (Gaussian Splatting) SLAM…
cs.CV2025
Ov3R: Open-Vocabulary Semantic 3D Reconstruction from RGB Videos
Ziren Gong, Xiaohan Li, Fabio Tosi +4
We present Ov3R, a novel framework for open-vocabulary semantic 3D reconstruction from RGB video streams, designed to advance Spatial AI. The system features two key components: CL…