12 citations · 13 across the 2 of their papers we have counts for
3 papers
cs.CV2024★ 12 cited
Towards Deconfounded Image-Text Matching with Causal Inference
Wenhui Li, Xinqi Su, Dan Song +3
Prior image-text matching methods have shown remarkable performance on many benchmark datasets, but most of them overlook the bias in the dataset, which exists in intra-modal and i…
cs.CV2023★ 1 cited
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
Dan Song, Xinwei Fu, Ning Liu +5
Large-scale pre-trained models have demonstrated impressive performance in vision and language tasks within open-world scenarios. Due to the lack of comparable pre-trained models f…
cs.CV2016
Multi-Camera Action Dataset for Cross-Camera Action Recognition Benchmarking
Wenhui Li, Yongkang Wong, An-An Liu +3
Action recognition has received increasing attention from the computer vision and machine learning communities in the last decade. To enable the study of this problem, there exist…