activity
20142023
most citedControlVideo: Training-free Controllable Text-to-Video Generation

34 citations · 64 across the 14 of their papers we have counts for

collaborators
Showing cs.CVShow all

10 papers · 1 filter

cs.CV20235 cited

Multi-modal Prompting for Low-Shot Temporal Action Localization

Chen Ju, Zeqian Li, Peisen Zhao +5

In this paper, we consider the problem of temporal action localization under low-shot (zero-shot & few-shot) scenario, with the goal of detecting and classifying the action instanc…

cs.CV20233 cited

Lformer: Text-to-Image Generation with L-shape Block Parallel Decoding

Jiacheng Li, Longhui Wei, ZongYuan Zhan +4

Generative transformers have shown their superiority in synthesizing high-fidelity and high-resolution images, such as good diversity and training stability. However, they suffer f…

cs.CV2022

Dilated Context Integrated Network with Cross-Modal Consensus for Temporal Emotion Localization in Videos

Juncheng Li, Junlin Xie, Linchao Zhu +8

Understanding human emotions is a crucial ability for intelligent robots to provide better human-robot interactions. The existing works are limited to trimmed video-level emotion c…

cs.CV2022

SdAE: Self-distillated Masked Autoencoder

Yabo Chen, Yuchen Liu, Dongsheng Jiang +4

With the development of generative-based self-supervised learning (SSL) approaches like BeiT and MAE, how to learn good representations by masking random patches of the input image…

cs.CV20223 cited

Skeleton-Parted Graph Scattering Networks for 3D Human Motion Prediction

Maosen Li, Siheng Chen, Zijing Zhang +3

Graph convolutional network based methods that model the body-joints' relations, have recently shown great promise in 3D skeleton-based human motion prediction. However, these meth…

cs.CV2022

Active Pointly-Supervised Instance Segmentation

Chufeng Tang, Lingxi Xie, Gang Zhang +3

The requirement of expensive annotations is a major burden for training a well-performed instance segmentation model. In this paper, we present an economic active learning setting,…