8 papers
Parameter-Dynamic Adaptive Fusion and Calibration Network for RGBT Tracking
Zhaoding Ding, Chenglong Li, Jiandong Jin +2
Existing RGBT trackers typically employ fusion functions with fixed parameters across different targets and scenarios. Although dynamic-architecture methods improve fusion flexibil…
T2I-VeRW: Part-level Fine-grained Perception for Text-to-Image Vehicle Retrieval
Xiao Wang, Ziwen Wang, Weizhe Kong +5
Vehicle Re-identification (Re-ID) aims to retrieve the most similar image to a given query from images captured by non-overlapping cameras. Extending vehicle Re-ID from image-only…
Vehicle-centric Perception via Multimodal Structured Pre-training
Wentao Wu, Xiao Wang, Chenglong Li +2
Vehicle-centric perception plays a crucial role in many intelligent systems, including large-scale surveillance systems, intelligent transportation, and autonomous driving. Existin…
Segment Any Vehicle: Semantic and Visual Context Driven SAM and A Benchmark
Xiao Wang, Ziwen Wang, Wentao Wu +4
With the rapid advancement of autonomous driving, vehicle perception, particularly detection and segmentation, has placed increasingly higher demands on algorithmic performance. Pr…
CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training Framework
Wentao Wu, Xiao Wang, Chenglong Li +4
Event cameras have attracted increasing attention in recent years due to their advantages in high dynamic range, high temporal resolution, low power consumption, and low latency. S…
FDDet: Frequency-Decoupling for Boundary Refinement in Temporal Action Detection
Xinnan Zhu, Yicheng Zhu, Tixin Chen +2
Temporal action detection aims to locate and classify actions in untrimmed videos. While recent works focus on designing powerful feature processors for pre-trained representations…