3 papers
cs.CV2025
M2H: Multi-Task Learning with Efficient Window-Based Cross-Task Attention for Monocular Spatial Perception
U. V. B. L Udugama, George Vosselman, Francesco Nex
Deploying real-time spatial perception on edge devices requires efficient multi-task models that leverage complementary task information while minimizing computational overhead. Th…
cs.CV2020
Self-supervised monocular depth estimation from oblique UAV videos
Logambal Madhuanand, Francesco Nex, Michael Ying Yang
UAVs have become an essential photogrammetric measurement as they are affordable, easily accessible and versatile. Aerial images captured from UAVs have applications in small and l…
cs.CV2020
Real-time Semantic Segmentation with Context Aggregation Network
Michael Ying Yang, Saumya Kumaar, Ye Lyu +1
With the increasing demand of autonomous systems, pixelwise semantic segmentation for visual scene understanding needs to be not only accurate but also efficient for potential real…