14 papers
MAGneT-3D: Monocular and Domain-Generalizable Temporal 3D Detection
Mohamed Kotb, Johannes Meier, Christoph Reich +3
Monocular temporal 3D detection aims to detect objects in 3D, given a monocular video. Query-based 3D detectors unify detection and cross-view association, but their learnable quer…
ECHO: Ego-Centric modeling of Human-Object interactions
Ilya A. Petrov, Vladimir Guzov, Riccardo Marin +5
Modeling human-object interactions (HOI) from an egocentric perspective is a critical yet challenging task, particularly when relying on sparse signals from wearable devices like s…
LeAD-M3D: Leveraging Asymmetric Distillation for Real-Time Monocular 3D Detection
Johannes Meier, Jonathan Michel, Oussema Dhaouadi +7
Real-time monocular 3D object detection remains challenging due to severe depth ambiguity, viewpoint shifts, and the high computational cost of 3D reasoning. Existing approaches ei…
SemCityLoc: Aerial 6DoF Localization Using Semantic 3D City Models
Jingfeng Mao, Xuyang Chen, Qilin Zhang +6
Aerial 6DoF localization typically relies on precise GNSS signals or radiometrically rich 3D reconstructions, limiting scalability and on-board deployment. We propose SemCityLoc, a…
IDEAL-M3D: Instance Diversity-Enriched Active Learning for Monocular 3D Detection
Johannes Meier, Florian Günther, Riccardo Marin +3
Monocular 3D detection relies on just a single camera and is therefore easy to deploy. Yet, achieving reliable 3D understanding from monocular images requires substantial annotatio…
GrounDiff: Diffusion-Based Ground Surface Generation from Digital Surface Models
Oussema Dhaouadi, Johannes Meier, Jacques Kaiser +1
Digital Terrain Models (DTMs) represent the bare-earth elevation and are important in numerous geospatial applications. Such data models cannot be directly measured by sensors and…