5 papers
COVTrack++: Learning Open-Vocabulary Multi-Object Tracking from Continuous Videos via a Synergistic Paradigm
Zekun Qian, Wei Feng, Ruize Han +1
Multi-Object Tracking (MOT) has traditionally focused on a few specific categories, restricting its applicability to real-world scenarios involving diverse objects. Open-Vocabulary…
Language-Assisted Image Clustering Guided by Discriminative Relational Signals and Adaptive Semantic Centers
Jun Ma, Xu Zhang, Zhengxing Jiao +4
Language-Assisted Image Clustering (LAIC) augments the input images with additional texts with the help of vision-language models (VLMs) to promote clustering performance. Despite…
PVNet: Point-Voxel Interaction LiDAR Scene Upsampling Via Diffusion Models
Xianjing Cheng, Lintai Wu, Zuowen Wang +3
Accurate 3D scene understanding in outdoor environments heavily relies on high-quality point clouds. However, LiDAR-scanned data often suffer from extreme sparsity, severely hinder…
Unsupervised 3D Point Cloud Completion via Multi-view Adversarial Learning
Lintai Wu, Xianjing Cheng, Yong Xu +2
In real-world scenarios, scanned point clouds are often incomplete due to occlusion issues. The tasks of self-supervised and weakly-supervised point cloud completion involve recons…
Synthetic-To-Real Video Person Re-ID
Xiangqun Zhang, Wei Feng, Ruize Han +3
Person re-identification (Re-ID) is an important task and has significant applications for public security and information forensics, which has progressed rapidly with the developm…