28 papers
Mask What Matters: Saliency-Guided Video Self-Supervised Learning for Autonomous Driving
Christopher Lang, Alexander Braun, Abhinav Valada
Video self-supervised learning through masked spatiotemporal prediction has emerged as a promising paradigm for learning feature representations from unlabeled data. However, exist…
Spotted: Location-informed Reidentification of Hyenas and Leopards in Camera Trap Surveys
Halil Sina Kelebek, Julia Hindel, Kobus Hoffman +9
Animal re-identification (ReID) in camera-trap surveys remains challenging due to low image quality, strong variation in illumination and viewpoint, and highly imbalanced numbers o…
Streaming Gaussian Encoding for 4D Panoptic Occupancy Tracking
Maximilian Luz, Thomas Nürnberg, Yakov Miron +1
Camera-based 4D panoptic occupancy tracking (4D-POT) is a promising paradigm for holistic scene understanding from multi-view imagery, enabling joint reasoning about geometry, sema…
Latent Gaussian Splatting for 4D Panoptic Occupancy Tracking
Maximilian Luz, Rohit Mohan, Thomas Nürnberg +3
Capturing 4D spatiotemporal scene structure is crucial for the safe and reliable operation of robots in dynamic environments. However, existing approaches typically address only pa…
Joint Target-Less Intrinsic and Extrinsic Camera-LiDAR Calibration using Deep Point Correspondences
Simon Bultmann, Daniele Cattaneo, Abhinav Valada
Accurate camera-LiDAR calibration is a prerequisite for robust multi-modal perception in robotics. Recent target-less approaches based on deep point correspondences achieve remarka…
AnchorD: Metric Grounding of Monocular Depth Using Factor Graphs
Simon Dorer, Martin Büchner, Nick Heppert +1
Dense and accurate depth estimation is essential for robotic manipulation, grasping, and navigation, yet currently available depth sensors are prone to errors on transparent, specu…