papers

Publications (15)

cs.CV2026

DriveVA: Video Action Models are Zero-Shot Drivers

Mengmeng Liu, Diankun Zhang, Jiuming Liu +7

cs.CV2025

4DSTR: Advancing Generative 4D Gaussians with Spatial-Temporal Rectification for High-Quality and Consistent 4D Generation

Mengmeng Liu, Jiuming Liu, Yunpeng Zhang +4

cs.RO2026

Mono-Hydra++: Real-Time Monocular Scene Graph Construction with Multi-Task Learning for 3D Indoor Mapping

U. V. B. L. Udugama, George Vosselman, Francesco Nex

cs.CV2026

A Comparison of Multi-View Stereo Methods for Photogrammetric 3D Reconstruction: From Traditional to Learning-Based Approaches

Yawen Li, George Vosselman, Francesco Nex

cs.CV2020

Self-supervised monocular depth estimation from oblique UAV videos

Logambal Madhuanand, Francesco Nex, Michael Ying Yang

cs.CV2025

Real-Time Bundle Adjustment for Ultra-High-Resolution UAV Imagery Using Adaptive Patch-Based Feature Tracking

Selim Ahmet Iz, Francesco Nex, Norman Kerle +2

cs.CV2026

M2H-MX: Multi-Task Semantic and Geometric Perception for Real-Time Monocular 3D Scene Graph Construction

U. V. B. L. Udugama, George Vosselman, Francesco Nex

cs.CV2025

TopoLiDM: Topology-Aware LiDAR Diffusion Models for Interpretable and Realistic LiDAR Point Cloud Generation

Jiuming Liu, Zheng Huang, Mengmeng Liu +4

cs.CV2021

Real-time Semantic Segmentation with Context Aggregation Network

Michael Ying Yang, Saumya Kumaar, Ye Lyu +1

cs.CV2023

Lite-Mono: A Lightweight CNN and Transformer Architecture for Self-Supervised Monocular Depth Estimation

Ning Zhang, Francesco Nex, George Vosselman +1

cs.CV2026

UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation

Mengmeng Liu, Diankun Zhang, Jiuming Liu +7

cs.CV2026

Semantic-Driven Scale and Spatial Selection for Efficient Cross-Modal Alignment in Referring Remote Sensing Image Segmentation

Kun Li, Shengxi Gui, Francesco Nex +1

cs.CV2026

ZeD-MAP: Bundle Adjustment Guided Zero-Shot Depth Maps for Real-Time Aerial Imaging

Selim Ahmet Iz, Francesco Nex, Norman Kerle +2

cs.CV2025

M2H: Multi-Task Learning with Efficient Window-Based Cross-Task Attention for Monocular Spatial Perception

U. V. B. L Udugama, George Vosselman, Francesco Nex

cs.RO2023

Channel-Aware Distillation Transformer for Depth Estimation on Nano Drones

Ning Zhang, Francesco Nex, George Vosselman +1