15 papers · 1 filter
@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology
Xin Jiang, Junwei Zheng, Ruiping Liu +4
As Vision-Language Models (VLMs) advance, human-centered Assistive Technologies (ATs) for helping People with Visual Impairments (PVIs) are evolving into generalists, capable of pe…
Occlusion-Aware Seamless Segmentation
Yihong Cao, Jiaming Zhang, Hao Shi +5
Panoramic images can broaden the Field of View (FoV), occlusion-aware prediction can deepen the understanding of the scene, and domain adaptation can transfer across viewing domain…
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping
Jiale Wei, Junwei Zheng, Ruiping Liu +3
In the field of autonomous driving, Bird's-Eye-View (BEV) perception has attracted increasing attention in the community since it provides more comprehensive information compared w…
CoBEV: Elevating Roadside 3D Object Detection with Depth and Height Complementarity
Hao Shi, Chengshan Pang, Jiaming Zhang +6
Roadside camera-driven 3D object detection is a crucial task in intelligent transportation systems, which extends the perception range beyond the limitations of vision-centric vehi…
OAFuser: Towards Omni-Aperture Fusion for Light Field Semantic Segmentation
Fei Teng, Jiaming Zhang, Kunyu Peng +3
Light field cameras are capable of capturing intricate angular and spatial details. This allows for acquiring complex light patterns and details from multiple angles, significantly…
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation
Ruiping Liu, Kailun Yang, Alina Roitberg +5
Semantic segmentation benchmarks in the realm of autonomous driving are dominated by large pre-trained transformers, yet their widespread adoption is impeded by substantial computa…