activity
20172023
most citedBinary Generative Adversarial Networks for Image Retrieval

84 citations · 273 across the 31 of their papers we have counts for

collaborators

39 papers

cs.CV2023

MotionZero:Exploiting Motion Priors for Zero-shot Text-to-Video Generation

Sitong Su, Litao Guo, Lianli Gao +2

Zero-shot Text-to-Video synthesis generates videos based on prompts without any videos. Without motion information from videos, motion priors implied in prompts are vital guidance.…

cs.CV20235 cited

CUCL: Codebook for Unsupervised Continual Learning

Chen Cheng, Jingkuan Song, Xiaosu Zhu +3

The focus of this study is on Unsupervised Continual Learning (UCL), as it presents an alternative to Supervised Continual Learning which needs high-quality manual labeled data. Th…

cs.CV20236 cited

Continual Referring Expression Comprehension via Dual Modular Memorization

Heng Tao Shen, Cheng Chen, Peng Wang +3

Referring Expression Comprehension (REC) aims to localize an image region of a given object described by a natural-language expression. While promising performance has been demonst…

cs.CV202314 cited

Class Gradient Projection For Continual Learning

Cheng Chen, Ji Zhang, Jingkuan Song +1

Catastrophic forgetting is one of the most critical challenges in Continual Learning (CL). Recent approaches tackle this problem by projecting the gradient update orthogonal to the…

cs.CV20231 cited

MovieFactory: Automatic Movie Creation from Text using Large Generative Models for Language and Images

Junchen Zhu, Huan Yang, Huiguo He +6

In this paper, we present MovieFactory, a powerful framework to generate cinematic-picture (30721280), film-style (multi-scene), and multi-modality (sounding) movies on the…

cs.CV2022

Visual Commonsense-aware Representation Network for Video Captioning

Pengpeng Zeng, Haonan Zhang, Lianli Gao +3

Generating consecutive descriptions for videos, i.e., Video Captioning, requires taking full advantage of visual representation along with the generation process. Existing video ca…