10 citations · 14 across the 5 of their papers we have counts for
5 papers
Knowledge Distillation from Single to Multi Labels: an Empirical Study
Youcai Zhang, Yuzhuo Qin, Hengwei Liu +3
Knowledge distillation (KD) has been extensively studied in single-label image classification. However, its efficacy for multi-label classification remains relatively unexplored. I…
Dense RGB SLAM with Neural Implicit Maps
Heng Li, Xiaodong Gu, Weihao Yuan +3
There is an emerging trend of using neural implicit functions for map representation in Simultaneous Localization and Mapping (SLAM). Some pioneer works have achieved encouraging r…
Vision-Language Matching for Text-to-Image Synthesis via Generative Adversarial Networks
Qingrong Cheng, Keyu Wen, Xiaodong Gu
Text-to-image synthesis aims to generate a photo-realistic and semantic consistent image from a specific text description. The images synthesized by off-the-shelf models usually co…
A Unified Two-Stage Group Semantics Propagation and Contrastive Learning Network for Co-Saliency Detection
Zhenshan Tan, Cheng Chen, Keyu Wen +2
Co-saliency detection (CoSOD) aims at discovering the repetitive salient objects from multiple images. Two primary challenges are group semantics extraction and noise object suppre…
Contrastive Cross-Modal Knowledge Sharing Pre-training for Vision-Language Representation Learning and Retrieval
Keyu Wen, Zhenshan Tan, Qingrong Cheng +2
Recently, the cross-modal pre-training task has been a hotspot because of its wide application in various down-streaming researches including retrieval, captioning, question answer…