activity
20242026
collaborators
Showing 2024Show all

15 papers · 1 filter

cs.CV2024

@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology

Xin Jiang, Junwei Zheng, Ruiping Liu +4

As Vision-Language Models (VLMs) advance, human-centered Assistive Technologies (ATs) for helping People with Visual Impairments (PVIs) are evolving into generalists, capable of pe…

cs.CV2024

Occlusion-Aware Seamless Segmentation

Yihong Cao, Jiaming Zhang, Hao Shi +5

Panoramic images can broaden the Field of View (FoV), occlusion-aware prediction can deepen the understanding of the scene, and domain adaptation can transfer across viewing domain…

cs.CV2024

OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping

Jiale Wei, Junwei Zheng, Ruiping Liu +3

In the field of autonomous driving, Bird's-Eye-View (BEV) perception has attracted increasing attention in the community since it provides more comprehensive information compared w…

cs.CV2024

CoBEV: Elevating Roadside 3D Object Detection with Depth and Height Complementarity

Hao Shi, Chengshan Pang, Jiaming Zhang +6

Roadside camera-driven 3D object detection is a crucial task in intelligent transportation systems, which extends the perception range beyond the limitations of vision-centric vehi…

cs.CV2024

OAFuser: Towards Omni-Aperture Fusion for Light Field Semantic Segmentation

Fei Teng, Jiaming Zhang, Kunyu Peng +3

Light field cameras are capable of capturing intricate angular and spatial details. This allows for acquiring complex light patterns and details from multiple angles, significantly…

cs.CV2024

TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation

Ruiping Liu, Kailun Yang, Alina Roitberg +5

Semantic segmentation benchmarks in the realm of autonomous driving are dominated by large pre-trained transformers, yet their widespread adoption is impeded by substantial computa…