4 papers · 1 filter
ReasonVQA: A Multi-hop Reasoning Benchmark with Structural Knowledge for Visual Question Answering
Duong T. Tran, Trung-Kien Tran, Manfred Hauswirth +1
In this paper, we propose a new dataset, ReasonVQA, for the Visual Question Answering (VQA) task. Our dataset is automatically integrated with structured encyclopedic knowledge and…
Collaborative Perceiver: Elevating Vision-based 3D Object Detection via Local Density-Aware Spatial Occupancy
Jicheng Yuan, Manh Nguyen Duc, Qian Liu +2
Vision-based bird's-eye-view (BEV) 3D object detection has advanced significantly in autonomous driving by offering cost-effectiveness and rich contextual information. However, exi…
Cooperative Students: Navigating Unsupervised Domain Adaptation in Nighttime Object Detection
Jicheng Yuan, Anh Le-Tuan, Manfred Hauswirth +1
Unsupervised Domain Adaptation (UDA) has shown significant advancements in object detection under well-lit conditions; however, its performance degrades notably in low-visibility s…
Co-Learning: Towards Semi-Supervised Object Detection with Road-side Cameras
Jicheng Yuan, Anh Le-Tuan, Ali Ganbarov +2
Recently, deep learning has experienced rapid expansion, contributing significantly to the progress of supervised learning methodologies. However, acquiring labeled data in real-wo…