activity
20172024
most citedBinary Generative Adversarial Networks for Image Retrieval

84 citations · 313 across the 32 of their papers we have counts for

collaborators

42 papers

cs.CV20235 cited

CUCL: Codebook for Unsupervised Continual Learning

Chen Cheng, Jingkuan Song, Xiaosu Zhu +3

The focus of this study is on Unsupervised Continual Learning (UCL), as it presents an alternative to Supervised Continual Learning which needs high-quality manual labeled data. Th…

cs.CV20236 cited

Continual Referring Expression Comprehension via Dual Modular Memorization

Heng Tao Shen, Cheng Chen, Peng Wang +3

Referring Expression Comprehension (REC) aims to localize an image region of a given object described by a natural-language expression. While promising performance has been demonst…

cs.CV202314 cited

Class Gradient Projection For Continual Learning

Cheng Chen, Ji Zhang, Jingkuan Song +1

Catastrophic forgetting is one of the most critical challenges in Continual Learning (CL). Recent approaches tackle this problem by projecting the gradient update orthogonal to the…

cs.CV20231 cited

MovieFactory: Automatic Movie Creation from Text using Large Generative Models for Language and Images

Junchen Zhu, Huan Yang, Huiguo He +6

In this paper, we present MovieFactory, a powerful framework to generate cinematic-picture (30721280), film-style (multi-scene), and multi-modality (sounding) movies on the…

cs.CV2022

Progressive Tree-Structured Prototype Network for End-to-End Image Captioning

Pengpeng Zeng, Jinkuan Zhu, Jingkuan Song +1

Studies of image captioning are shifting towards a trend of a fully end-to-end paradigm by leveraging powerful visual pre-trained models and transformer-based generation architectu…

cs.CV2022

Learning Dual-Fused Modality-Aware Representations for RGBD Tracking

Shang Gao, Jinyu Yang, Zhe Li +3

With the development of depth sensors in recent years, RGBD object tracking has received significant attention. Compared with the traditional RGB object tracking, the addition of t…