30 citations · 30 across the 5 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023
Enhanced Multimodal Representation Learning with Cross-modal KD
Mengxi Chen, Linyu Xing, Yu Wang +1
This paper explores the tasks of leveraging auxiliary modalities which are only available at training to enhance multimodal representation learning through cross-modal Knowledge Di…
cs.CV2022
Self-Supervised Masking for Unsupervised Anomaly Detection and Localization
Chaoqin Huang, Qinwei Xu, Yanfeng Wang +2
Recently, anomaly detection and localization in multimedia data have received significant attention among the machine learning community. In real-world applications such as medical…