activity
20142024
most citedConvolutional Network for Attribute-driven and Identity-preserving Human Face Generation

48 citations · 144 across the 42 of their papers we have counts for

collaborators
Showing cs.CVShow all

35 papers · 1 filter

cs.CV2024

Arbitrary-Scale Video Super-Resolution with Structural and Textural Priors

Wei Shang, Dongwei Ren, Wanying Zhang +3

Arbitrary-scale video super-resolution (AVSR) aims to enhance the resolution of video frames, potentially at various scaling factors, which presents several challenges regarding sp…

cs.CV2024

Multi-modal Crowd Counting via a Broker Modality

Haoliang Meng, Xiaopeng Hong, Chenhao Wang +2

Multi-modal crowd counting involves estimating crowd density from both visual and thermal/depth images. This task is challenging due to the significant gap between these distinct m…

cs.CV20243 cited

Evaluation of Text-to-Video Generation Models: A Dynamics Perspective

Mingxiang Liao, Hannan Lu, Xinyu Zhang +6

Comprehensive and constructive evaluation protocols play an important role in the development of sophisticated text-to-video (T2V) generation models. Existing evaluation protocols…

cs.CV2024

SmartControl: Enhancing ControlNet for Handling Rough Visual Conditions

Xiaoyu Liu, Yuxiang Wei, Ming Liu +4

Human visual imagination usually begins with analogies or rough sketches. For example, given an image with a girl playing guitar before a building, one may analogously imagine how…

cs.CV2024

Responsible Visual Editing

Minheng Ni, Yeli Shen, Lei Zhang +1

With recent advancements in visual synthesis, there is a growing risk of encountering images with detrimental effects, such as hate, discrimination, or privacy violations. The rese…

cs.CV20244 cited

A Comprehensive Survey on 3D Content Generation

Jian Liu, Xiaoshui Huang, Tianyu Huang +8

Recent years have witnessed remarkable advances in artificial intelligence generated content(AIGC), with diverse input modalities, e.g., text, image, video, audio and 3D. The 3D is…