activity
20242026
collaborators
Showing 2025Show all

7 papers · 1 filter

cs.CV2025

Vehicle-centric Perception via Multimodal Structured Pre-training

Wentao Wu, Xiao Wang, Chenglong Li +2

Vehicle-centric perception plays a crucial role in many intelligent systems, including large-scale surveillance systems, intelligent transportation, and autonomous driving. Existin…

cs.CV2025

ICPL-ReID: Identity-Conditional Prompt Learning for Multi-Spectral Object Re-Identification

Shihao Li, Chenglong Li, Aihua Zheng +2

Multi-spectral object re-identification (ReID) brings a new perception perspective for smart city and intelligent transportation applications, effectively addressing challenges fro…

cs.CV2025

CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training Framework

Wentao Wu, Xiao Wang, Chenglong Li +4

Event cameras have attracted increasing attention in recent years due to their advantages in high dynamic range, high temporal resolution, low power consumption, and low latency. S…

cs.CV2025

Breaking Shallow Limits: Task-Driven Pixel Fusion for Gap-free RGBT Tracking

Andong Lu, Yuanzhi Guo, Wanyu Wang +3

Current RGBT tracking methods often overlook the impact of fusion location on mitigating modality gap, which is key factor to effective tracking. Our analysis reveals that shallowe…

cs.CV2025

Towards General Multimodal Visual Tracking

Andong Lu, Mai Wen, Jinhu Wang +4

Existing multimodal tracking studies focus on bi-modal scenarios such as RGB-Thermal, RGB-Event, and RGB-Language. Although promising tracking performance is achieved through lever…

cs.CV2025

Large Language Model Guided Progressive Feature Alignment for Multimodal UAV Object Detection

Wentao Wu, Chenglong Li, Xiao Wang +2

Existing multimodal UAV object detection methods often overlook the impact of semantic gaps between modalities, which makes it difficult to achieve accurate semantic and spatial al…