5 papers
T2I-VeRW: Part-level Fine-grained Perception for Text-to-Image Vehicle Retrieval
Xiao Wang, Ziwen Wang, Weizhe Kong +5
Vehicle Re-identification (Re-ID) aims to retrieve the most similar image to a given query from images captured by non-overlapping cameras. Extending vehicle Re-ID from image-only…
Vehicle-centric Perception via Multimodal Structured Pre-training
Wentao Wu, Xiao Wang, Chenglong Li +2
Vehicle-centric perception plays a crucial role in many intelligent systems, including large-scale surveillance systems, intelligent transportation, and autonomous driving. Existin…
Segment Any Vehicle: Semantic and Visual Context Driven SAM and A Benchmark
Xiao Wang, Ziwen Wang, Wentao Wu +4
With the rapid advancement of autonomous driving, vehicle perception, particularly detection and segmentation, has placed increasingly higher demands on algorithmic performance. Pr…
DehazeMamba: SAR-guided Optical Remote Sensing Image Dehazing with Adaptive State Space Model
Zhicheng Zhao, Jinquan Yan, Chenglong Li +2
Optical remote sensing image dehazing presents significant challenges due to its extensive spatial scale and highly non-uniform haze distribution, which traditional single-image de…
Large Language Model Guided Progressive Feature Alignment for Multimodal UAV Object Detection
Wentao Wu, Chenglong Li, Xiao Wang +2
Existing multimodal UAV object detection methods often overlook the impact of semantic gaps between modalities, which makes it difficult to achieve accurate semantic and spatial al…