papers
Publications (15)
cs.CV2026
DriveVA: Video Action Models are Zero-Shot Drivers
Mengmeng Liu, Diankun Zhang, Jiuming Liu +7
cs.CV2025
4DSTR: Advancing Generative 4D Gaussians with Spatial-Temporal Rectification for High-Quality and Consistent 4D Generation
Mengmeng Liu, Jiuming Liu, Yunpeng Zhang +4
cs.RO2026
Mono-Hydra++: Real-Time Monocular Scene Graph Construction with Multi-Task Learning for 3D Indoor Mapping
U. V. B. L. Udugama, George Vosselman, Francesco Nex
cs.CV2026
A Comparison of Multi-View Stereo Methods for Photogrammetric 3D Reconstruction: From Traditional to Learning-Based Approaches
Yawen Li, George Vosselman, Francesco Nex
cs.CV2020
Self-supervised monocular depth estimation from oblique UAV videos
Logambal Madhuanand, Francesco Nex, Michael Ying Yang
cs.CV2025
Real-Time Bundle Adjustment for Ultra-High-Resolution UAV Imagery Using Adaptive Patch-Based Feature Tracking
Selim Ahmet Iz, Francesco Nex, Norman Kerle +2
cs.CV2026
M2H-MX: Multi-Task Semantic and Geometric Perception for Real-Time Monocular 3D Scene Graph Construction
U. V. B. L. Udugama, George Vosselman, Francesco Nex
cs.CV2025
TopoLiDM: Topology-Aware LiDAR Diffusion Models for Interpretable and Realistic LiDAR Point Cloud Generation
Jiuming Liu, Zheng Huang, Mengmeng Liu +4
cs.CV2021
Real-time Semantic Segmentation with Context Aggregation Network
Michael Ying Yang, Saumya Kumaar, Ye Lyu +1
cs.CV2023
Lite-Mono: A Lightweight CNN and Transformer Architecture for Self-Supervised Monocular Depth Estimation
Ning Zhang, Francesco Nex, George Vosselman +1
cs.CV2026
UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation
Mengmeng Liu, Diankun Zhang, Jiuming Liu +7
cs.CV2026
Semantic-Driven Scale and Spatial Selection for Efficient Cross-Modal Alignment in Referring Remote Sensing Image Segmentation
Kun Li, Shengxi Gui, Francesco Nex +1
cs.CV2026
ZeD-MAP: Bundle Adjustment Guided Zero-Shot Depth Maps for Real-Time Aerial Imaging
Selim Ahmet Iz, Francesco Nex, Norman Kerle +2
cs.CV2025
M2H: Multi-Task Learning with Efficient Window-Based Cross-Task Attention for Monocular Spatial Perception
U. V. B. L Udugama, George Vosselman, Francesco Nex
cs.RO2023
Channel-Aware Distillation Transformer for Depth Estimation on Nano Drones
Ning Zhang, Francesco Nex, George Vosselman +1