10 papers
Koala-36M: A Large-scale Video Dataset Improving Consistency between Fine-grained Conditions and Video Content
Qiuheng Wang, Yukai Shi, Jiarong Ou +10
With the continuous progress of visual generation technologies, the scale of video datasets has grown exponentially. The quality of these datasets plays a pivotal role in the perfo…
Owl-1: Omni World Model for Consistent Long Video Generation
Yuanhui Huang, Wenzhao Zheng, Yuan Gao +5
Video generation models (VGMs) have received extensive attention recently and serve as promising candidates for general-purpose large vision models. While they can only generate sh…
Motion Inversion for Video Customization
Luozhou Wang, Ziyang Mai, Guibao Shen +6
In this work, we present a novel approach for motion customization in video generation, addressing the widespread gap in the exploration of motion representation within video gener…
VideoTetris: Towards Compositional Text-to-Video Generation
Ye Tian, Ling Yang, Haotian Yang +9
Diffusion models have demonstrated great success in text-to-video (T2V) generation. However, existing methods may face challenges when handling complex (long) video generation scen…
Towards Unified 3D Hair Reconstruction from Single-View Portraits
Yujian Zheng, Yuda Qiu, Leyang Jin +5
Single-view 3D hair reconstruction is challenging, due to the wide range of shape variations among diverse hairstyles. Current state-of-the-art methods are specialized in recoverin…
ViMo: Generating Motions from Casual Videos
Liangdong Qiu, Chengxing Yu, Yanran Li +6
Although humans have the innate ability to imagine multiple possible actions from videos, it remains an extraordinary challenge for computers due to the intricate camera movements…