collaborators

10 papers

cs.CV2025

Koala-36M: A Large-scale Video Dataset Improving Consistency between Fine-grained Conditions and Video Content

Qiuheng Wang, Yukai Shi, Jiarong Ou +10

With the continuous progress of visual generation technologies, the scale of video datasets has grown exponentially. The quality of these datasets plays a pivotal role in the perfo…

cs.CV2024

Owl-1: Omni World Model for Consistent Long Video Generation

Yuanhui Huang, Wenzhao Zheng, Yuan Gao +5

Video generation models (VGMs) have received extensive attention recently and serve as promising candidates for general-purpose large vision models. While they can only generate sh…

cs.CV2024

Motion Inversion for Video Customization

Luozhou Wang, Ziyang Mai, Guibao Shen +6

In this work, we present a novel approach for motion customization in video generation, addressing the widespread gap in the exploration of motion representation within video gener…

cs.CV2024

VideoTetris: Towards Compositional Text-to-Video Generation

Ye Tian, Ling Yang, Haotian Yang +9

Diffusion models have demonstrated great success in text-to-video (T2V) generation. However, existing methods may face challenges when handling complex (long) video generation scen…

cs.CV2024

Towards Unified 3D Hair Reconstruction from Single-View Portraits

Yujian Zheng, Yuda Qiu, Leyang Jin +5

Single-view 3D hair reconstruction is challenging, due to the wide range of shape variations among diverse hairstyles. Current state-of-the-art methods are specialized in recoverin…

cs.CV2024

ViMo: Generating Motions from Casual Videos

Liangdong Qiu, Chengxing Yu, Yanran Li +6

Although humans have the innate ability to imagine multiple possible actions from videos, it remains an extraordinary challenge for computers due to the intricate camera movements…