11 citations · 24 across the 5 of their papers we have counts for
5 papers
MagicProp: Diffusion-based Video Editing via Motion-aware Appearance Propagation
Hanshu Yan, Jun Hao Liew, Long Mai +2
This paper addresses the issue of modifying the visual appearance of videos while preserving their motion. A novel framework, named MagicProp, is proposed, which disentangles the v…
MagicEdit: High-Fidelity and Temporally Coherent Video Editing
Jun Hao Liew, Hanshu Yan, Jianfeng Zhang +2
In this report, we present MagicEdit, a surprisingly simple yet effective solution to the text-guided video editing task. We found that high-fidelity and temporally coherent video-…
MagicAvatar: Multimodal Avatar Generation and Animation
Jianfeng Zhang, Hanshu Yan, Zhongcong Xu +2
This report presents MagicAvatar, a framework for multimodal video generation and animation of human avatars. Unlike most existing methods that generate avatar-centric videos direc…
Delving Deeper into Data Scaling in Masked Image Modeling
Cheng-Ze Lu, Xiaojie Jin, Qibin Hou +3
Understanding whether self-supervised learning methods can scale with unlimited data is crucial for training large-scale models. In this work, we conduct an empirical study on the…
Associating Spatially-Consistent Grouping with Text-supervised Semantic Segmentation
Yabo Zhang, Zihao Wang, Jun Hao Liew +4
In this work, we investigate performing semantic segmentation solely through the training on image-sentence pairs. Due to the lack of dense annotations, existing text-supervised me…