collaborators

8 papers

cs.CV2026

Parameter-Dynamic Adaptive Fusion and Calibration Network for RGBT Tracking

Zhaoding Ding, Chenglong Li, Jiandong Jin +2

Existing RGBT trackers typically employ fusion functions with fixed parameters across different targets and scenarios. Although dynamic-architecture methods improve fusion flexibil…

cs.CV2026

T2I-VeRW: Part-level Fine-grained Perception for Text-to-Image Vehicle Retrieval

Xiao Wang, Ziwen Wang, Weizhe Kong +5

Vehicle Re-identification (Re-ID) aims to retrieve the most similar image to a given query from images captured by non-overlapping cameras. Extending vehicle Re-ID from image-only…

cs.CV2025

Vehicle-centric Perception via Multimodal Structured Pre-training

Wentao Wu, Xiao Wang, Chenglong Li +2

Vehicle-centric perception plays a crucial role in many intelligent systems, including large-scale surveillance systems, intelligent transportation, and autonomous driving. Existin…

cs.CV2025

Segment Any Vehicle: Semantic and Visual Context Driven SAM and A Benchmark

Xiao Wang, Ziwen Wang, Wentao Wu +4

With the rapid advancement of autonomous driving, vehicle perception, particularly detection and segmentation, has placed increasingly higher demands on algorithmic performance. Pr…

cs.CV2025

CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training Framework

Wentao Wu, Xiao Wang, Chenglong Li +4

Event cameras have attracted increasing attention in recent years due to their advantages in high dynamic range, high temporal resolution, low power consumption, and low latency. S…

cs.CV2025

FDDet: Frequency-Decoupling for Boundary Refinement in Temporal Action Detection

Xinnan Zhu, Yicheng Zhu, Tixin Chen +2

Temporal action detection aims to locate and classify actions in untrimmed videos. While recent works focus on designing powerful feature processors for pre-trained representations…