50 citations · 240 across the 31 of their papers we have counts for
32 papers · 1 filter
X-modaler: A Versatile and High-performance Codebase for Cross-modal Analytics
Yehao Li, Yingwei Pan, Jingwen Chen +2
With the rise and development of deep learning over the past decade, there has been a steady momentum of innovation and breakthroughs that convincingly push the state-of-the-art of…
A Low Rank Promoting Prior for Unsupervised Contrastive Learning
Yu Wang, Jingyang Lin, Qi Cai +4
Unsupervised learning is just at a tipping point where it could really take off. Among these approaches, contrastive learning has seen tremendous progress and led to state-of-the-a…
Contextual Transformer Networks for Visual Recognition
Yehao Li, Ting Yao, Yingwei Pan +1
Transformer with self-attention has led to the revolutionizing of natural language processing field, and recently inspires the emergence of Transformer-style architecture design wi…
Scheduled Sampling in Vision-Language Pretraining with Decoupled Encoder-Decoder Network
Yehao Li, Yingwei Pan, Ting Yao +2
Despite having impressive vision-language (VL) pretraining with BERT-based encoder for VL understanding, the pretraining of a universal encoder-decoder for both VL understanding an…
Joint Contrastive Learning with Infinite Possibilities
Qi Cai, Yu Wang, Yingwei Pan +2
This paper explores useful modifications of the recent development in contrastive learning via novel probabilistic modeling. We derive a particular form of contrastive loss named J…
SeCo: Exploring Sequence Supervision for Unsupervised Representation Learning
Ting Yao, Yiheng Zhang, Zhaofan Qiu +2
A steady momentum of innovations and breakthroughs has convincingly pushed the limits of unsupervised image representation learning. Compared to static 2D images, video has one mor…