565 citations · 1.4k across the 55 of their papers we have counts for
4 papers · 2 filters
Contrastive Visual-Linguistic Pretraining
Lei Shi, Kai Shuang, Shijie Geng +6
Several multi-modality representation learning approaches such as LXMERT and ViLBERT have been proposed recently. Such approaches can achieve superior performance due to the high-l…
Character Matters: Video Story Understanding with Character-Aware Relations
Shijie Geng, Ji Zhang, Zuohui Fu +3
Different from short videos and GIFs, video stories contain clear plots and lists of principal characters. Without identifying the connection between appearing people and character…
TIME: Text and Image Mutual-Translation Adversarial Networks
Bingchen Liu, Kunpeng Song, Yizhe Zhu +2
Focusing on text-to-image (T2I) generation, we propose Text and Image Mutual-Translation Adversarial Networks (TIME), a lightweight but effective model that jointly learns a T2I ge…
Long Short-Term Sample Distillation
Liang Jiang, Zujie Wen, Zhongping Liang +7
In the past decade, there has been substantial progress at training increasingly deep neural networks. Recent advances within the teacher--student training paradigm have establishe…