59 citations · 131 across the 34 of their papers we have counts for
55 papers
SA4Depth: Consistent Pose-Depth Scale Alignment for Self-Supervised Monocular Depth Estimation
Changxuan Li, Nadine Berner, Nassir Navab +2
Self-supervised depth estimation from monocular sequences relies on the joint learning of a depth and a pose network. Despite abundant research done to improve the depth network, e…
Language-Guided Open-World Anomaly Segmentation
Klara Reichard, Nikolas Brasch, Nassir Navab +1
Open-world and anomaly segmentation methods seek to enable autonomous driving systems to detect and segment both known and unknown objects in real-world scenes. However, existing m…
SING3R-SLAM: Submap-based Indoor Monocular Gaussian SLAM with 3D Reconstruction Priors
Kunyi Li, Michael Niemeyer, Sen Wang +3
Recent advances in dense 3D reconstruction have demonstrated strong capability in accurately capturing local geometry. However, extending these methods to incremental global recons…
Epipolar Geometry Improves Video Generation Models
Orest Kupyn, Théo Uscidda, Marta Tintore Gazulla +3
Video generation models have advanced significantly through the latent diffusion transformers trained with rectified flow techniques. Yet these models still struggle with geometric…
GALA: Guided Attention with Language Alignment for Open Vocabulary Gaussian Splatting
Elena Alegret, Kunyi Li, Sen Wang +5
3D scene reconstruction and understanding have gained increasing popularity, yet existing methods still struggle to capture fine-grained, language-aware 3D representations from 2D…
Self-supervised Latent Space Optimization with Nebula Variational Coding
Yida Wang, David Joseph Tan, Nassir Navab +1
Deep learning approaches process data in a layer-by-layer way with intermediate (or latent) features. We aim at designing a general solution to optimize the latent manifolds to imp…