activity
20242026
collaborators

7 papers

cs.CV2026

ReasonVQA: A Multi-hop Reasoning Benchmark with Structural Knowledge for Visual Question Answering

Duong T. Tran, Trung-Kien Tran, Manfred Hauswirth +1

In this paper, we propose a new dataset, ReasonVQA, for the Visual Question Answering (VQA) task. Our dataset is automatically integrated with structured encyclopedic knowledge and…

cs.CV2025

Collaborative Perceiver: Elevating Vision-based 3D Object Detection via Local Density-Aware Spatial Occupancy

Jicheng Yuan, Manh Nguyen Duc, Qian Liu +2

Vision-based bird's-eye-view (BEV) 3D object detection has advanced significantly in autonomous driving by offering cost-effectiveness and rich contextual information. However, exi…

cs.CV2025

Cooperative Students: Navigating Unsupervised Domain Adaptation in Nighttime Object Detection

Jicheng Yuan, Anh Le-Tuan, Manfred Hauswirth +1

Unsupervised Domain Adaptation (UDA) has shown significant advancements in object detection under well-lit conditions; however, its performance degrades notably in low-visibility s…

cs.RO2024

Performance Evaluation of ROS2-DDS middleware implementations facilitating Cooperative Driving in Autonomous Vehicle

Sumit Paul, Danh Lephuoc, Manfred Hauswirth

In the autonomous vehicle and self-driving paradigm, cooperative perception or exchanging sensor information among vehicles over wireless communication has added a new dimension. G…

cs.CV2024

Co-Learning: Towards Semi-Supervised Object Detection with Road-side Cameras

Jicheng Yuan, Anh Le-Tuan, Ali Ganbarov +2

Recently, deep learning has experienced rapid expansion, contributing significantly to the progress of supervised learning methodologies. However, acquiring labeled data in real-wo…

cs.RO2024

A comparison of extended object tracking with multi-modal sensors in indoor environment

Jiangtao Shuai, Martin Baerveldt, Manh Nguyen-Duc +3

This paper presents a preliminary study of an efficient object tracking approach, comparing the performance of two different 3D point cloud sensory sources: LiDAR and stereo camera…